# Capable Enough and Runs Anywhere
Four uncoordinated releases today — Google DeepMind's Gemma 4 QAT checkpoints, NVIDIA's Nemotron 3.5 ASR, Moonshot AI's Kimi Code CLI, and the Thousand Token Wood multi-agent experiment — converge on a single infrastructure thesis: the unit of value delivery is shifting from one large hosted model call to composed systems of smaller, locally-runnable, format-reliable components. The episode argues this is real progress on narrow sub-problems, while the hard capability problems remain untouched. The efficiency narrative is too flattering; this is infrastructure plumbing, not the autonomous-agent breakthrough the coverage implies.
## Thread 1: The Efficiency Offensive
## Thread 2: Open-Source as Infrastructure Wedge
## Thread 3: Small Models as Economic Actors
## Cross-Story / Counter-Narrative Sources