News
Every entry here is generated from the same changelog that drives the
Signal Room report itself
(data/market-state.json) – model roster additions, benchmark refreshes,
and fixes, in the order they happened.
Added Moonshot's official Kimi K3 API pricing: $3/M uncached input, $0.30/M cached...
Tuesday, July 21, 2026 in News
Model roster · Added Moonshot’s official Kimi K3 API pricing: $3/M uncached input, $0.30/M cached input, and $15/M output; the 30/70 workload blend is $11.40/M before reasoning-effort effects.
Added Apertus-v1.1-4B-Instruct, the largest newly released Apertus Mini checkpoint...
Tuesday, July 21, 2026 in News
Model roster · Added Apertus-v1.1-4B-Instruct, the largest newly released Apertus Mini checkpoint: fully open Apache 2.0 weights and data, 4K context, 1.7T-token distillation, 1,811 languages, and official BF16, FP8, NVFP4A16, INT3, INT4 and INT6 …
Updated the action queue and recommended routing: Fable for hardest retained-data...
Saturday, July 18, 2026 in News
Routing · Updated the action queue and recommended routing: Fable for hardest retained-data workloads, Terra for default engineering, Luna for high-volume subagents, with explicit escalation rules.
Restored GPT-5.5 family visibility, verified Haiku 4.5 as the latest public Haiku...
Saturday, July 18, 2026 in News
Data · Restored GPT-5.5 family visibility, verified Haiku 4.5 as the latest public Haiku, replaced quality compound display with quality vs selected reference, and removed non-actionable headline cost/policy counters.
Added source-grounded GPT-5.6 Sol, Terra and Luna pricing, 1.05M context, benchmark...
Saturday, July 18, 2026 in News
Data · Added source-grounded GPT-5.6 Sol, Terra and Luna pricing, 1.05M context, benchmark registers and selectable reference configurations. Added documented quality, speed, cost and capability composites in data/report-metrics.json.
Added Kimi K3 from Moonshot primary sources: 2.8T sparse MoE, 1M context, native...
Saturday, July 18, 2026 in News
Model roster · Added Kimi K3 from Moonshot primary sources: 2.8T sparse MoE, 1M context, native vision, max-only thinking at launch, API availability, and a vendor-suite quality comparison against Fable 5.
Restored GPT-5.3-Codex-Spark (Feb 12, 2026 release; ChatGPT Pro research preview, 128K...
Saturday, June 06, 2026 in News
Fix · Restored GPT-5.3-Codex-Spark (Feb 12, 2026 release; ChatGPT Pro research preview, 128K context, 1000+ tok/s on Cerebras) and Hermes Agent v0.16.0 (Nous Research, MIT, self-hosted multi-platform agent) — both were incorrectly removed in v1.5.0 …
Dashboard market sweep v1.5.0: real-world re-grounding
Saturday, June 06, 2026 in News
Policy · Dashboard market sweep v1.5.0: real-world re-grounding. Replaced fictional Mythos/GPT-5.5-Cyber rows with verified models; added Nvidia Nemotron coalition, Kimi K2.6, GLM-5, Cohere Command A+, SubQ 1M-Preview.
Nvidia releases Nemotron 3 Ultra (550B/55B MoE, hybrid Mamba-Transformer, 1M context...
Thursday, June 04, 2026 in News
Model roster · Nvidia releases Nemotron 3 Ultra (550B/55B MoE, hybrid Mamba-Transformer, 1M context, NVIDIA Open Model License) at Computex — first frontier-scale open model from Nvidia.
Nvidia Nemotron Coalition formed: Black Forest Labs, Cursor, LangChain, Mistral...
Thursday, June 04, 2026 in News
Model roster · Nvidia Nemotron Coalition formed: Black Forest Labs, Cursor, LangChain, Mistral, Perplexity, Reflection AI, Sarvam, Thinking Machines Lab as inaugural members.