GPT-4.1 GPT-4.1 is OpenAI's latest and most advanced flagship model, significantly improving upon GPT-4 Turbo in performance across benchmarks, speed, and cost-effectiveness.
Benchmark results Service providers 90.2%
87.9%
87.4%
87.3%
74.8%
72.2%
72.0%
70.8%
68.0%
66.3%
65.8%
65.5%
61.7%
58.0%
57.2%
56.7%
54.6%
52.9%
51.6%
49.4%
49.1%
48.1%
46.4%
46.3%
46.2%
38.3%
28.9%
25.0%
19.0%
5.4%
Pricing, uptime, and speed via OpenRouter — updated Jul 17, 2026, 04:19 AM.
Provider Status Input Output Limits Uptime Speed Notes OpenAI available $2.00/Mtok cache $0.50/Mtok $8.00/Mtok 1.0M tokens context 33K tokens max output 99.3% 5m 98% 549 ms p50 TTFT 20 tok/s p50 $0.01/web search Azure available $2.20/Mtok cache $0.55/Mtok $8.80/Mtok 1.0M tokens context 33K tokens max output — — cache $0.01/web search