- November 2025
-
Code Execution Cuts MCP Agent Token Costs (anthropic.com)
- June 2025
-
What Anthropic Learned Building a Multi-Agent Researcher (anthropic.com)
- April 2025
-
Qwen3 Puts Reasoning on a Switch (qwenlm.github.io)
-
AI as Normal Technology (normaltech.ai)
- March 2025
-
OLMo 2 32B: Fully Open Catches Up to Closed (allenai.org)
- January 2025
-
DeepSeek-R1: An Open Model Matches a Closed Reasoner (huggingface.co)
- November 2024
-
Reward Hacking: Why Better Models Game You More (lilianweng.github.io)
-
Tülu 3 Opens Up the Post-Training Recipe (allenai.org)
- September 2024
-
OpenAI o1 and the Start of Test-Time Reasoning (openai.com)
- August 2024
-
Can AI Scaling Continue Through 2030? (epoch.ai)