Which models lead, where API prices moved, and what it costs to run them. Measured from live data.
Tracking 753 hosted models across OpenRouter, ppq.ai, and kie.ai.
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI GPT-5 | $0.02 | $0.20 |
| OpenAI o-series | $0.55 | $2.20 |
| Anthropic Claude Opus | $1.43 | $7.15 |
| Anthropic Claude Sonnet | $0.85 | $4.28 |
| Google Gemini 2.5 Pro | $0.38 | $3.00 |
| Google Gemini Flash | $0.05 | $0.20 |
Gemma 3 27B IT: cumulative cost over 12 months
The API saves about $48 over a year at this workload.
Local vs API assumes 50M tokens a month, split evenly between input and output, with the cheapest cloud GPU that fits the model billed at 8 hours a day.
The most-starred open agent frameworks, with how their stars and downloads moved.
New papers posted to arXiv in the categories that drive applied AI. A rough gauge of how fast the field is moving, not a quality measure.
10,599 papers in total. Papers land in more than one category, so the four numbers above add up to more than the real total.
A point-in-time read of major provider uptime, taken when this report was generated.
One email the day each edition is out, plus The Agent Roundup, our weekly note for AI builders. Free, unsubscribe any time.
