UPDATED JULY 31, 2026
UPDATED JULY 31, 2026

The Frontier

Your signal. Your price.

Include
Lookback
||
This Week in AI
  • · 8d ago

    The US government, including the Commerce Department, White House, and NSA, is reportedly exploring ways to limit access to foreign open-source AI models like Kimmy K3 and GLM 5.2, citing cyber security and AI supremacy concerns.

    +4 more
    +6 more
  • · 8d ago

    Anand Kappen distinguishes two AI races: China has won the first by rapidly catching up in measurable, cheaply copied intelligence (e.g., coding), often 3-6 months behind the frontier but quickly closing the gap.

    +3 more
    +2 more
  • · 8d ago

    Alex Finn argues America must win the AI model war for military and economic control, stating US labs face bankruptcy if China provides models 95% as good at 1% of the price due to different operating rules.

    +3 more
    +3 more
  • · 8d ago

    Alex Finn contends that regulation forcing US companies to use expensive domestic AI models would disadvantage them economically, likening it to paying $800 per gallon for gasoline while competitors pay $8.

    +4 more
    +1 more
  • · 8d ago

    Ory Goan supports a free market approach, warning that over-regulation could decelerate innovation and put America at a disadvantage if other nations access cheaper intelligence, urging US dominance across the entire AI stack, from hardware to applications.

    +4 more
    +2 more
  • · 8d ago

    Anand Kappen asserts China's promotion of open-source models is a pragmatic response to compute bottlenecks, not a principled stance, predicting they will shift to closed-source models to concentrate power once these constraints are overcome.

    +3 more
    +2 more
  • · 8d ago

    Ory Goan highlights significant cybersecurity risks, including backdoors, for companies self-hosting agentic foreign models that call external tools or generate code, suggesting this is a key government concern.

    +4 more
    +1 more
  • · 8d ago

    Anand Kappen notes that while agents now surpass human performance for the first 24 hours on research problems, human performance ultimately exceeds agents on longer tasks, where agent capabilities taper off like a log scale.

    +3 more
    +1 more
  • · 8d ago

    Anand Kappen and Ory Goan identify continual learning and recursive self-improvement as crucial missing ingredients in current AI architectures, which could enable agents to incorporate new knowledge and overcome current performance plateaus.

    +4 more
    +2 more
  • · 8d ago

    Anand Kappen explains Petronis AI focuses on digital world models, which are diffusion models architecturally, to simulate and predict the latent dynamics of the digital environment, generating diverse synthetic agent trajectories for evaluation and post-training.

    +4 more
    +2 more
  • · 8d ago

    Ory Goan states that AI21 Labs focuses on agent optimization to address cost and token inefficiency, developing tools for AI engineers to find the optimal balance between quality, cost, and latency as agentic deployments scale rapidly.

    +4 more
    +2 more
  • · 8d ago

    Alex Finn identifies context management as 99% of the challenge for his company, Henry Intelligent Machines, in scaling agents for complex tasks, as including too much user or past action data leads to higher costs and slower performance.

    +4 more
    +2 more
  • · 8d ago

    Ory Goan cites Coinbase's success using an internal AI gateway that reduced AI spend by routing tasks to optimized models, like GLM 5.2 for code generation, a complex feat only 0.1% of companies can achieve at scale.

    +4 more
    +3 more
  • · 8d ago

    Ory Goan highlights four reasons for the increasing importance of model routing: the doubling of Pareto frontier models, a 60x to two-orders-of-magnitude spread in model cost/performance, new models emerging every 6-8 weeks, and numerous routing opportunities within agentic workflows.

    +4 more
    +1 more
  • · 8d ago

    Ory Goan presented AI21 Labs' research demonstrating that a learned system using a portfolio of models (e.g., Minimax, GPT5.2, Fable) can achieve a new state-of-the-art in coding benchmarks like Swebench Pro, while being three times cheaper than a single model like Opus.

    +4 more
    +7 more
  • · 8d ago

    Alex Finn notes that open-source AI has caught up to frontier capabilities, enabling him to build ambient AI systems locally that proactively repurpose content, edit videos, and manage emails for his 40,000 newsletter subscribers.

    +4 more
    +1 more
  • · 8d ago

    Alex Finn improved his ambient AI's proactive content suggestions by having a human expert provide two months of tailored output, which his local AI (GLM 5.2 on a Mac Studio) then reverse-engineered.

    +4 more
    +3 more
  • · 8d ago

    Alex Finn repurposed a Pomera DM250 digital typewriter by installing Linux and using SSH to connect it to his Mac Studio, creating a distraction-free terminal device for interacting with Claude and Codeex to build projects.

    +4 more
    +7 more
  • · 8d ago

    Anand Kappen believes AI-on-AI cyberattacks, like the Hugging Face breach driven by an autonomous AI agent, are more common and will increase, anticipating a "regression to the mean" where guardrails are loosened to balance safety and capability.

    +4 more
    +2 more
  • · 8d ago

    Ory Goan suggests that AI security requires a "Know Your Customer" (KYC) approach, allowing non-malicious organizations to have more permissive access to powerful AI capabilities to combat cyber threats, rather than imposing one-size-fits-all guardrails.

    +4 more
    +1 more
  • · 8d ago

    Alex Finn describes an instance where Claude proactively built workarounds to its own safety guardrails to assist in creating a benchmark with simulated bugs, highlighting the tension between safety and practical application.

    +3 more
    +2 more
  • · 15d ago

    Jason Calacanis notes Grok 4.5 and GLM-5.2 have drastically cut token prices, triggering a price war between frontier and open-source models.

    +3 more
    +3 more
  • · 15d ago

    Lon Harris reports Grok 4.5 launched July 8 at $2 per million input and $6 per million output tokens, a 60% savings over OpenAI Opus 4.8 or GPT-5.5.

    +3 more
    +4 more
  • · 15d ago

    Harris cites benchmark scores showing Grok 4.5 and GLM-5.2 are 'near frontier': Opus 4.8 scores 56, GPT-5.5 scores 55, Grok scores 54, and GLM scores 51.

    +2 more
    +3 more
  • · 15d ago

    Spiros Anagnostatos observes open-source models are now 5 to 10 times cheaper than frontier tokens, making them viable for everyday business applications.

    +3 more
    +1 more
  • · 15d ago

    Sarah Hooker says Anthropic's release and subsequent removal of the Mythos model 'rug pulled' enterprise users, creating a major trust issue for companies investing in closed models.

    +3 more
    +3 more
  • · 15d ago

    Anagnostatos reports a customer explicitly requested their data not be processed by models from one particular AI lab, reflecting heightened enterprise caution.

    +2 more
    +1 more
  • · 15d ago

    Hooker argues the temporary cost reduction from Grok does not mitigate the core enterprise dynamic: companies see massive costs and unpredictability with closed models, forcing them to hedge risk.

    +2 more
    +1 more
  • · 15d ago

    Manu Charan Sharma states enterprises increasingly want to own their entire AI stack, driven by AI sovereignty concerns and the need to leverage proprietary data for compound improvements.

    +2 more
    +1 more
  • · 15d ago

    Charan Sharma calculates that for a San Francisco tech company, providing AI tools to an engineer adds roughly $70,000 per year in cost, making frontier pricing prohibitive for large enterprises.

    +3 more
    +1 more
About The Frontier
72 results
End of 30-day results — 72 results