AI Summary
5 min readKimi K3: A New Frontier for Open Source AI
Chinese AI company Moonshot AI has released Kimi K3, a 2.8 trillion parameter model that its benchmark scores suggest rivals—and in some areas exceeds—the best proprietary models from OpenAI and Anthropic. The model is open source, with full weights scheduled for release by July 27th, and available now for free at kimi.com with no credit card required. The release lands alongside news that Google has fallen months behind schedule on its Gemini 3.5 Pro flagship, and that Major League Baseball has banned dugout iPads from accessing generative AI for in-game decisions after discovering roughly a third of teams were already using it that way.
Kimi K3 Closes the Open Source Gap
The headline numbers are striking. Kimi K3 scored 88.3 on Terminal Bench 2.1, trailing only GPT 5.6 Sol's 88.8. On GDP Val Double A V2, a benchmark measuring real-world tasks across 44 occupations and nine industries, K3 scored 1687—third behind Claude Fable 5 Max at 1815 and GPT 5.6 Sol Max at 1747.8, but ahead of Claude Opus 4.8 at 1600. On BrowseComp, a benchmark for long-horizon information seeking, K3 achieved a state-of-the-art score of 91.2 out of 100, accomplished in a single-agent setup using its 1 million token context window without any context compression or additional management techniques.
Continue reading the full summary in the app — free to try.
Read Full Summary →Free • No credit card required
Never miss an episode of Tech Brew Ride Home
Get every new episode summarized in your inbox — free, ~5 minutes to read.
No spam. Unsubscribe anytime.
What you'll learn
- 1 (01:04) **Moonshot AI Releases Kimi K3** - A 2.8 trillion parameter open-source model that rivals Claude Opus 4.8 and GPT 5.5, with full model weights to be released by July 27th.
- 2 (03:55) **Kimi K3 Benchmark Performance** - Scores third on GDP Val Double A V2 (1687), behind Claude Fable 5 Max and GPT 5.6 Sol Max, and ahead of Claude Opus 4.8.
- 3 (05:36) **Kimi K3 Autonomous Agent Demonstration** - Over 48 hours, the model autonomously designed a functional chip to run a nano version of itself, completing the full construction pipeline from architecture to verification.
- 4 (08:25) **Google Delays Gemini 3.5 Pro** - Alphabet is reportedly months behind schedule on its most powerful flagship AI model, causing internal frustration.
- 5 (12:06) **MLB Bans AI on Dugout iPads** - The league has prohibited the use of league-provided iPads for generative AI in in-game strategy, with at least a third of teams reportedly using custom apps for substitutions and pitch calling.
- 6 (15:18) **Weekend Long Read: Siri AI Preview** - A recommendation to try the new iOS public beta and Siri AI, which is changing how users interact with their iPhones.
+ Full timestamped outline available in the app
Show Notes
Moonshot AI released Kimi K3, a 2.8T-parameter model it says rivals Opus 4.8 and GPT-5.5. Google fell months behind on Gemini 3.5 Pro, MLB banned dugout iPads from accessing GenAI for in-game calls, and The Verge tested Siri AI.
- Moonshot AI releases Kimi K3, a 2.8T-parameter AI model that it says rivals Claude Opus 4.8 and GPT-5.5, and plans to release its full model weights by July 27 (VentureBeat)
- Sources: Google is months behind schedule on delivering Gemini 3.5 Pro as it tries to improve its capabilities, particularly in coding; GOOG closes down 4.43% (Bloomberg)
- Memo: MLB bans the use of league-provided dugout iPads to access GenAI for in-game strategy calls; sources say at least a third of teams used AI this way (The Athletic)
Longreads
Subscribe to the ad-free feed.
Learn more about your ad choices. Visit megaphone.fm/adchoices
More from this podcast
Tech Brew Ride Home →