Key Takeaways
- Average latency for GPT-4o on OpenRouter is 250ms
- P99 latency under 2 seconds for Claude 3.5 Sonnet
- Throughput of 1,200 tokens/second for Mixtral 8x22B
- Cost savings of up to 40% on Llama 3.1 compared to direct providers via OpenRouter
- OpenRouter generated $5M+ in provider payouts in 2024 YTD
- Average spend per user $25/month
- OpenRouter supports over 200 AI models from 20+ providers as of Q3 2024
- OpenRouter routes to 15+ inference engines including vLLM and TensorRT-LLM
- 50+ open-source models available with fallbacks
- OpenRouter processed more than 500 million tokens in a single day peak in September 2024
- Daily API requests exceeded 10 million in August 2024
- Peak concurrent requests hit 50,000 per minute
- OpenRouter has 150,000+ active monthly users
- 75% user retention rate month-over-month
- 1.2 million API keys issued since launch
OpenRouter delivers sub second median latency, 99.99 percent uptime, and up to 40 percent lower costs.
Related reading
01 · Category
API Performance21 stats
API Performance Interpretation
02 · Category
Economic Impact21 stats
Economic Impact Interpretation
03 · Category
Model Diversity21 stats
Model Diversity Interpretation
More related reading
04 · Category
Usage Volume19 stats
Usage Volume Interpretation
05 · Category
User Adoption22 stats
User Adoption Interpretation
Cite This Report
This report is designed to be cited. We maintain stable URLs and versioned verification dates. Copy the format appropriate for your publication below.
Helena Kowalczyk. (2026, February 24). OpenRouter Statistics. Gitnux. https://gitnux.org/openrouter-statistics
Helena Kowalczyk. "OpenRouter Statistics." Gitnux, 24 Feb 2026, https://gitnux.org/openrouter-statistics.
Helena Kowalczyk. 2026. "OpenRouter Statistics." Gitnux. https://gitnux.org/openrouter-statistics.
Sources & references
6 datasets cited across this report · attribution is report-level

