/
© 2026 RiffOn. All rights reserved.

Get your free personalized podcast brief

We scan new podcasts and send you the top 5 insights daily.

  1. Super Data Science: ML & AI Podcast with Jon Krohn
  2. 1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier
1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

Super Data Science: ML & AI Podcast with Jon Krohn · Jul 24, 2026

China's Kimi K3, a 2.8T parameter open-weight model, disrupts the AI frontier with aggressive pricing and competitive performance, challenging US labs.

Moonshot AI's 2.8T-Parameter Kimi K3 Is Feasible by Activating Only 2% of Its Experts Per Token

The massive 2.8 trillion parameter count of Kimi K3 is misleading for cost analysis. Its Mixture of Experts (MOE) architecture activates only 16 of its 896 expert submodules per token. This makes the model computationally efficient and affordable for inference despite its enormous total capacity.

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier thumbnail

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

Super Data Science: ML & AI Podcast with Jon Krohn·3 days ago

Open-Weight Models Threaten to Turn Frontier AI From a Premium Service into a Commodity

By promising to release its model weights, Moonshot's Kimi K3 offers enterprises frontier-class AI on their own infrastructure, eliminating per-token fees and data privacy concerns. This combination of low cost, high performance, and customer control directly challenges the premium, tightly controlled service model of Western AI labs like OpenAI and Anthropic.

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier thumbnail

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

Super Data Science: ML & AI Podcast with Jon Krohn·3 days ago

Chinese AI Labs Bypass US Chip Sanctions With Architectural Innovations, Not Just More GPUs

Despite facing U.S. export controls on advanced chips, Moonshot AI's Kimi K3 demonstrates that significant performance gains are achievable through architectural innovations. Novel techniques like "Kimi Delta Attention" and "attention residuals" delivered a 2.5x scaling efficiency improvement, proving that software and model design can circumvent hardware limitations.

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier thumbnail

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

Super Data Science: ML & AI Podcast with Jon Krohn·3 days ago

Kimi K3's "Always-On Reasoning" Creates Hidden Costs by Billing Unseen Thinking as Premium Output Tokens

A key operational detail of Kimi K3 is its locked "always-on" reasoning mode. The model consumes tokens for internal "thinking" processes, and these are billed at the expensive output rate of $15 per million. This makes it powerful for complex tasks but potentially wasteful and costly for simple lookups.

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier thumbnail

1012: The Open-Weight 2.8-Trillion Parameter Competing at the Frontier

Super Data Science: ML & AI Podcast with Jon Krohn·3 days ago