What Happened
This week’s developments point to three decisions for AI teams: when parallel agents justify their cost, how to evaluate AI for scientific work, and what controls may be needed beyond voluntary commitments. Import AI reports that a four-agent swarm matched a single agent’s performance in half the time while using roughly twice the…
What Happened
Agent capability, efficiency and deployment infrastructure were the main themes of the week. OpenAI introduced several models and products, including GPT-6.1 Sol and Astra Ultrafast. Google announced Gemini 4 Argon with a one-million-token output limit and a company-reported 77.9% score on DeepSWE v1.1. NVIDIA launched an Open Agent Safety Platform, while Strands Agents…
What Happened
This week’s AI coverage concentrated on cheaper models, agent tooling and a warning about benchmarks. OpenAI launched GPT-6.1 Sol at $2 per million input tokens and $10 per million output tokens; reported arena placements put Sol Max at fifth in Agent Arena and Gemini 4 Argon High at first in Text Arena. OpenAI…
What Happened
Airbnb described an “inside-out AI” strategy: improve how its teams build software, then carry those capabilities into the guest experience. It says AI authors 60% of its code and reports roughly 1.6 times as many pull requests per engineer. Its Everest context graph is intended to help engineers navigate specialist code and reuse…
What Happened
Frontier-model announcements drew attention, but the more actionable developments concern how models are used and verified. Google DeepMind’s Gemini 4 Argon reportedly supports up to one million output tokens and leads many published benchmarks; some results are disputed, and access is limited to government users and trusted cyber defenders. OpenAI introduced persistent agents,…
What Happened
Three converging trends dominated the week: rapid model product moves and price competition, a spate of serious safety/tool‑use incidents, and platform innovations that change how agents are hosted and billed.
Model releases and pricing: Anthropic released Opus 5.5 and Sonnet 5.5 (lower cost/latency) while OpenAI countered with GPT‑6 Sol and GPT‑6.1…
Findings [1] 2026-09-29 The Sequence Knowledge - Issue 941: Learning RSI: The Model Is Frozen. The System Is Not. Here is an experience you have probably had by now. You set up an agent in the spring. Same model all quarter, weights frozen solid, not a single gradient step. And yet by summer the…
Findings [1] 2026-09-28 Import AI 474: Platonic mindspace; TPUs in space; Zhipu starts an outer RSI loop Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe.Subscribe nowAre minds patterns from a Platonic space, with bodies and…
Findings [1] 2026-09-27 The Sequence Radar - Issue 940: Last Week in AI: Opus 5.5 Gets Leaner, Meta Goes Wearable, Washington Talks to Beijing, and Claude Explores DNA Next Week in The Sequence:Another installment of our series about recursive self-improvement. We dive into Opus 5.5, DeepSeek’s amazing new paper about environments and Anthropic’s DNA…
Findings [1] 2026-09-25 OpenRouter: from Seed to Stripe — with OpenRouter’s Alex Atallah & AMP’s Anjney Midha From the earliest days of open-weight models to becoming the neutral routing layer for more than 10 million developers, OpenRouter is one of the clearest bets that the future of AI will be multi-model. In this episode,…
Findings [1] 2026-09-24 Foundries vs Navigators: Lowering the Cost of Science What does the future of science look like in the world of AI? Anthropic has some lofty goals for science and is even opening a wet lab. Meanwhile a quiet transformation1 is happening all across AI x Science.In this guest… A data portal…
Findings [1] 2026-09-23 🔬Bio-security is an AI Arms Race - Eric Nguyen (CEO, Radical Numerics) The OpenAI → Hugging Face attack has people asking “what else do we need to worry about?” and Anthropic’s filters flag two things: cyber-security and biology. The natural question is: what about bio-security, then? Clem Delangue argues that cyber-warfare…