What Happened
Recent signals show open-weight models are closing the performance gap with proprietary frontiers while new tooling and policy proposals accelerate capability diffusion and scrutiny. Evaluations report GLM‑5.2 near Claude Opus on narrow cyber tests and DeepSeek V4‑Pro positioned between Opus and GPT‑5; a long‑horizon test still shows a modest gap, but defenders have…
What Happened
Last week’s industry signals show a clear shift from monolithic scale toward openness, sparsity and extreme model compression, plus renewed focus on automated safety testing and governance. Key developments: Inkling (975B MoE, ~41B active, multimodal, 1M‑token context) was open‑sourced under Apache‑2.0; Moonshot announced a 2.8T Kimi K3 that activates a tiny fraction of…
What Happened
Market and research attention this week concentrated on one clear narrative shift: the community moved from a pure “compute moat” story to an efficiency stack thesis — i.e., gains from routing (MoE), quantization, data curation and kernel/perf engineering now matter as much as raw FLOPs [1].
Key signals driving that shift:
…
Executive Summary
This week saw major model releases and ecosystem moves that push open models toward frontier capabilities while increasing infrastructure and safety demands. Thinking Machines Lab published Inkling with day‑0 Apache‑2.0 weights and a fine‑tuning ecosystem, Meta announced Muse Spark 1.1 and related compute ambitions, and Moonshot released Kimi K3 (2.8T, 1M context) with…