Open AI Former CTO Mira's startup-general open source model Inkling Inkling: 975B, 41B total activations GLM-5.2: 744B, 40B total activations Supports 1M context and adjustable inference budget. Coding capabilities are not as good as GLM 5.2 Ultra Large MoE, but Active Parameters are controlled very low. The knowledge capacity continues to expand, but the cost of reasoning does not rise. This is very critical for long-running agents. Thinking Effort is adjustable. Different inference budgets are allocated to different tasks, and model capabilities begin to become a schedulable resource. Token Efficiency In the future, the competition will not only be accuracy, but also Quality/Token and Quality/Cost. Native multimodality + open weighting. Enterprises can fine-tune, deploy, not just call APIs.
English indexed field note
Open AI Former CTO Mira's startup-general open source model Inkling Inkling: 975B, 41B total activations
Open AI Former CTO Mira's startup-general open source model Inkling Inkling: 975B, 41B total activations GLM-5.2: 744B, 40B total activations Supports 1M context...