Price:

AI & TECH

Kimi k3 breaks US AI monopoly

Monday, July 20, 2026 · from 3 podcasts, 4 episodes
  • China’s Kimi K3 shattered the US-led AI duopoly, hitting top coding benchmarks on restricted hardware.
  • Open-weight models are now sovereign tools, forcing US allies to rethink dependency.
  • The model’s imminent open release triggers a global race for AI independence.

Kimi K3 didn’t just close the gap - it obliterated it. Moonshot AI’s 2.8 trillion parameter model leapt past Claude 3.5 Sonnet on coding benchmarks, a feat achieved without access to top-tier Nvidia chips. Instead, China used H-800s and domestic silicon from Huawei and Alibaba to sidestep US export controls.

This is a Sputnik moment with one key difference: the launch was silent. There was no fanfare, no global alert - just a model drop that rewrote the rules overnight. Emad Mostaque noted that China engineered around the compute wall, proving high-end AI no longer requires American permission. When the weights release in late July, any entity with enough GPUs can run it locally.

The geopolitical fallout is immediate. At the G7, Emmanuel Macron warned that the US now holds an AI kill switch. The UK’s request for an export carve-out was denied, souring European sentiment. Nathaniel Whittemore reported that trusted partner status no longer guarantees access to frontier models - a shift that’s accelerating sovereign AI development across Europe.

US labs now face a crisis of efficiency. While American companies spend billions on R&D, China leverages Western existence proofs to cut costs by 90%. Alex Weiser Gross pointed out that Kimi K3 uses standard transformer architecture - no secret sauce, just superior engineering and data quality. The result: 100x cost advantage.

Enterprises are adapting fast. Harvey and Microsoft now use compound architectures, pairing cheap open-weight models like DeepSeek V4 with expensive frontier models as advisors. This worker-advisor model cuts inference costs by 85% while improving performance - a direct response to the unpredictability of US export policy.

"The US government now controls the flow of intelligence."

- Nathaniel Whittemore, The AI Daily Brief

That control is backfiring. US companies are integrating Chinese models into core products to maintain cost efficiency. Microsoft is fine-tuning DeepSeek V4 for Copilot, creating a paradox: Washington bans Chinese access to US models while US firms embed Chinese AI into the American tech stack.

The shelf life of a top model is now weeks. Salim Ismail predicts daily frontier releases by January. In this environment, the value isn’t in the weights - it’s in the architecture that swaps them. The race is no longer about who builds the smartest model, but who can deploy and replace them fastest.

Kimi K3 didn’t end the duopoly - it ended the era of duopolies. The frontier is now a global free-for-all, where access, not innovation, determines power.

Source Intelligence

- Deep dive into what was said in the episodes

Urgent Update- AI Sputnik Moment: Kimi K3 Released w/ Emad Mostaque | Ep. 272Jul 19

  • China has prioritized humanoid robot development, with 150 companies, and stages public robot combat events that test demanding engineering aspects like balance, impact resistance, and locomotion.
Also from this episode: (15)

Models (11)

  • Moonshot AI, a Chinese lab, released Kimi K3, a multimodal, open-weight model with 2.8 trillion parameters, which rapidly achieved number one ranking in the Code Arena and six other domains.
  • Alex Weiser Gross notes Kimi models held state-of-the-art among open-weight models for 9 of the past 12 months, despite building on a recognizable transformer architecture without novel breakthroughs.
  • Imad Mostaque compares Kimi K3's engineering to Chinese EVs, achieving high performance and usability through manufacturing efficiency, despite using older hardware like H-800s under US export controls.
  • Dave Blundon describes Kimi K3's release as a "Sputnik times infinity" moment, enabling any entity to achieve near-frontier AI capabilities by using open-source weights and fine-tuning, independent of US models.
  • Salim Ismail argues frontier intelligence is a perishable asset with a shelf life of weeks, suggesting future value will reside in architectures that can rapidly swap and adapt different AI models.
  • Imad Mostaque reports Kimi K3 used the same compute as Inkling but achieved 2.5 times better data-to-intelligence conversion, optimized for Chinese silicon like Huawei 910 Ascend chips.
  • Since mid-April, 13 new frontier models have launched, averaging one every 10 days, a significant acceleration from previous years, with Alex Weiser Gross predicting daily releases by January.
  • PrismML's Bonside 27B is the first 27 billion parameter class model to run entirely on a smartphone, achieving ternary quantization that significantly reduces model size and increases speed fivefold.
  • Alex Weiser Gross predicts sub-1-bit quantization will go mainstream within the next year, and Imad Mostaque forecasts Kimi K3-level capability on a 16GB RAM MacBook by late next year due to distillation and new chipsets.
  • AI models are now statistically indistinguishable from human "superforecasters" in predicting novel events, implying cheap, tireless superhuman advisors for decisions in insurance, investing, and policy.
  • Salim Ismail argues that AI forecasting will make most senior management expertise obsolete, as AI systems can reproduce judgment without human biases, shifting focus to purpose and objectives.

Startups (1)

  • Alex Weiser Gross clarifies that Moonshot AI founder Yang Jilin started his Chinese startup, Recurrent AI, during his CMU PhD program, not after graduation due to visa issues.

Immigration (1)

  • Salim Ismail notes 70% of elite AI researchers are not U.S. citizens, with Chinese and Indian talent being prominent; he argues that the U.S. immigration system fails to retain this crucial talent.

Health (1)

  • Dr. Don Musilam reports that 3.3% of Fountain Life members, presumed healthy, have undetected cancer, emphasizing the critical role of early detection via tools like full body MRI for better cure rates.

AI Infrastructure (1)

  • Salim Ismail highlights that U.S. data centers consume 17 billion gallons of water annually, significantly less than golf courses (531 billion gallons) or California almond farming (1 trillion gallons).

The AI Duopoly Is Over: Grok 4. 5 , GPT-5 . 6 , and Muse Spark in One Week | #270Jul 13

  • OpenAI released GPT-5.6 on July 9th, a family of models including Sol, Tara, Luna, and Ultramode. The release moves OpenAI towards recursive self-improvement, using the high-end Sol model to post-train the lower-end Luna.
  • Elon Musk announced Grok 4.5 on July 8th. Meta released Muse Spark on July 9th, positioning it within its apps like WhatsApp and Messenger. Alex Gleas argues the frontier is no longer a duopoly; four American labs now operate at the optimal frontier.
  • Dave London says OpenAI's pivot from consumer to enterprise is evident; Chat GPT Work is a cargo-cult imitation of Anthropic's Claude app, focusing on revenue per token from code generation.
  • Peter Diamandis believes distribution is the new moat for AI models. Meta has 3.56 billion daily users, Google reaches 2.5 billion globally, and OpenAI has a billion monthly active users.
  • Alex Gleas sees an intelligence-polarized future: cheap, embedded AI on wearables versus high-cost frontier models in data centers driving scientific discovery. He argues profits from the high end fund the compute clusters.
  • A new EU law mandates infrared driver monitoring cameras in all new cars and vans sold from July 7th. The system triggers alerts if a driver glances away for more than 3.5 seconds above 31 mph. Brussels claims it will save 25,000 lives by 2038.
  • AI-generated performer Tilly Norwood was cast as the lead in the feature film 'Misaligned'. The Screen Actors Guild condemned the casting, arguing it devalues human artistry.
Also from this episode: (5)

Big Tech (1)

  • Apple sued OpenAI for trade secret theft, alleging OpenAI stole confidential files and code names to build its AI hardware with Johnny Ive. Salim Mayel thinks Apple filed in Northern California because it is desperate to slow competitors while catching up.

Space (2)

  • Elon Musk tweeted SpaceX will be worth more than Earth if it accomplishes its goals. Peter Diamandis notes Earth's material wealth is about $600 trillion; adding financial assets brings it to $1.7 quadrillion.
  • China's Long March 10B booster successfully landed on July 10th, a first for China. SpaceX's Falcon 9 has achieved 580 booster reflights; booster 1067 flew its 36th mission on July 11th.

Robotics (1)

  • 1X Neo unveiled a redesigned hand with 25 degrees of freedom, tendon-driven and waterproof. Burn Borneck plans to produce 10,000 units in 2026.

Safety (1)

  • Illinois Governor J.B. Pritzker signed SB 315, the Artificial Intelligence Safety Measures Act, requiring frontier labs earning over $500 million annually to conduct third-party safety audits and report incidents within 72 hours.

Kimi K3 and the Open-Weight Race, AI Treasure Hunting, Bitcoin Falls Further Behind GoldJul 17

  • Kimi K3 represents the biggest AI development of the last year: an open-source model with near-frontier performance that can be run sovereignly on owned hardware.
  • Open-source models like Kimi K3 will erode margins for frontier lab companies by offering comparable capability without paying for the model itself, only compute and electricity.
  • China's release of a competitive open-source model may be a geopolitical tactic to depress financing for US AI buildout before key IPOs, making it harder for American labs to raise capital.
  • DK uses AI agents paired with Strava heat maps, satellite imagery, and prompt engineering to research a multi-million dollar treasure hunt from the book 'There's Treasure Inside'.
  • Kevin Kelly argues latent spaces are infinite-dimensional worlds for human exploration; AI can tune concepts like 'Africa' or 'redness', but human volition is required to ask the questions.
  • AI agents today lack volition and get stuck in confirmation bias loops during research; they require human judgment to reset context and explore new hypothesis paths.
  • Current AI music models like Google's produce seven-out-of-ten beats but lag behind frontier models; IP restrictions prevent direct artist mashups, creating demand for open-source, permissionless alternatives.
Also from this episode: (6)

Open Source (2)

  • Project Loop, Spiral's AI security scanning tool, received a 5/5 usefulness rating from Bitcoin Core and positive feedback from eight initial projects.
  • The policy for Project Loop's free service should prioritize open-source public goods with high user impact; companies and pre-mined token projects present ethical and resource allocation challenges.

Trade (1)

  • China halted gold futures trading on the Shanghai exchange and continues aggressive gold accumulation while dumping US treasuries, signaling a move away from dollar reliance.

BTC Markets (1)

  • The Bitcoin-to-gold price ratio has fallen roughly 50-60% over the past year, reflecting Bitcoin's bear market and gold's price strength.

Protocol (1)

  • Stripe's potential acquisition of PayPal would likely converge PayPal's stablecoin onto Stripe's Tempo network and OpenUSD, buying users rather than integrating tech stacks.

Science (1)

  • Archaeological discoveries like Göbekli Tepe and LiDAR scans in the Amazon reveal civilizations far older and more extensive than previously believed, suggesting historical cycles of technological reset.

Is Kimi K3 Really Fable Class?Jul 17

  • Nathaniel Whittemore notes the G7 meeting featured unprecedented AI industry representation with leaders like Sam Altman, Demis Hassabis, Arthur Mensch, and Dario Amodei attending alongside heads of state.
  • Dario Amodei argued for international cooperation including structured access to frontier models, chip trade deals excluding China, and a unified approach to AI risks like cyberattacks and bioterrorism at the G7.
  • European leaders, like Emmanuel Macron, expressed concern that the US holds an AI 'kill switch' and pleaded for shared access to frontier models, fearing reliance on the US.
  • Sam Altman asserted AI regulation must be shaped by democratic institutions and society, not just corporations, and called for an international forum to establish global testing standards and risk analysis.
  • Whittemore reports the US did not concede on Fable access at the G7, and the UK's request for an export control carve-out was denied, souring the EU mood.
  • Noam Shazeer, co-author of the Transformer paper, left Google for OpenAI after Google spent $2.7 billion licensing his Character AI technology to retain him less than two years earlier.
  • OpenAI is sunsetting the Pulse daily briefing feature and expanding scheduled tasks to all paid subscribers, signaling a shift towards prioritizing users who need automated task workflows.
  • The Fable 5 ban has accelerated enterprise interest in open-source models for cost predictability and to avoid government kill switches, with media consensus pointing to open source as the biggest winner.
  • Chinese lab Moonshot AI released Kimi K 2.7 Code, claiming a 22% improvement on its code bench and 30% lower reasoning token usage compared to 2.6, though early users report it underwhelms in practice.
Also from this episode: (6)

AI Infrastructure (3)

  • Europe's AI sovereignty plan commits only 20 billion euros to build five gigafactories deploying around 100,000 GPUs, while US hyperscalers spend three times that monthly on AI data centers.
  • Open Router's Fusion API uses a panel of models routed by a judge and synthesizer to achieve frontier-level performance at half the price, validating a future where tasks are routed to specialized, cheaper models.
  • Harvey's experiment combining an open-weight GLM 5.1 worker with a closed frontier Opus 4.7 advisor increased performance and lowered costs, demonstrating that smart model routing beats using the most expensive model for every task.

Models (1)

  • The open-source model GLM 5.2 from ZAI beats GPT-5.5 and Opus 4.8 on some benchmarks at one-tenth the cost, fueling speculation about its distillation from Anthropic models and its viability as a Fable alternative.

Enterprise (1)

  • Microsoft is reportedly considering a locally hosted fine-tune of DeepSeek V4 for Copilot Co-work to offer cheaper AI access to enterprise customers, potentially normalizing Chinese models in US enterprise stacks.

Coding (1)

  • Cursor's Composer 2.5, built on a Kimi foundation, scores near Opus 4.7 and GPT-5.5 on benchmarks at a fraction of the cost, but user reports and updated agent-focused benchmarks show mixed real-world performance.