Refreshed current-source evidence: added OpenAI's GPT-5.6 launch table, Google's...

Benchmarks · Refreshed current-source evidence: added OpenAI’s GPT-5.6 launch table, Google’s Managed Agents update, and Scale’s public SWE-bench Pro leaderboard. Clarified that vendor launch tables and the public leaderboard are not directly comparable because their model versions and harnesses differ.