Key Takeaways
- Kimi K3 achieves near parity with Claude Fable 5 on software engineering tasks at ~35% the cost, but faces severe throughput issues on OpenRouter due to compute shortages.
- Qwen3.8 (2.4T params) preview is live, set to be open-sourced, positioning as one of the most powerful models.
- Open-weight models are defended as "AI communism" that prevents oligopoly, with strong support from Aravind Srinivas and Yann LeCun, while Deedy highlights compute as the new battleground.
- Meta's ~7GW compute by end-2026 could be the biggest wildcard in the AI infrastructure race.
- Ethan Mollick warns that AI capability curves remain steep and regulatory asymmetry (closed vs open) must be addressed.
1. Model Releases and Performance Comparisons
- Kimi K3 matches Claude Fable 5 on software engineering tasks at approximately 35% of the price, and performs even better at higher pass@k levels. However, on OpenRouter it struggles with throughput dropping from 30 to 13 tok/s, latency up to 72s, and time to first token over 20s, attributed to insufficient infrastructure ($500k for 8 B300s or $4M for a GB300 NVL72 rack). — via 1 2
- Qwen3.8 (2.4 trillion parameters) preview is now available and will be fully open-sourced. Aravind Srinivas calls it one of the most powerful models. — via 1
2. The Open vs Closed Model Debate and Compute Scarcity
- Aravind Srinivas describes a world dominated by open-weight models as "AI communism," arguing it is safer than narratives likening open weights to nuclear weapons. He notes that gradient updates allow progress and that expensive proprietary cloud AI (like Claude Fable 5 at $50/M output tokens) is unsustainable if cheaper open alternatives run on commodity hardware. — via 1 2
- Yann LeCun strongly endorses open-source AI, stating open weights prevent oligopoly, reduce costs, and bring economic and military advantages. He cites Linux and Android as successful precedents. — via 1 2 (multiple tweets condensed)
- Deedy emphasizes that the future of AI will be defined by compute availability. GPU providers demand 3-5 year commitments with 30% down; only hyperscalers, Grok, and Meta have locked-in deals. Meta with ~7GW compute by end-2026 could be the biggest wildcard. Despite model prices dropping (frontier from $60/M to $15/M over 3 years), demand has grown 3+ orders of magnitude and performance improved 32x, making compute access crucial. — via 1
- Ethan Mollick points out a regulatory asymmetry: the US is tightening oversight on frontier closed-source models while open-source models face little regulation, which must be resolved or will have major consequences. — via 1
3. Grok Build Updates and AI Infrastructure
- Elon Musk highlights the new Grok Build version, defaulting to Grok 4.5, with new commands and improvements. Grok Build is evolving into a full development environment integrating Vercel, Sentry, Stripe, emphasizing an end-to-end path from idea to deployment. — via 1 2
- Elon Musk asserts that faster progress toward higher intelligence will bring abundance and eventually eliminate conflict, and that future "builders" will number over 1 billion. — via 1 2
