Cheaper is left, more capable is up. 5 of 62 priced models are both cheaper and score at least as high. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.
Context Window in Context
Codestral256K
▸ o3200K
o3 Mini200K
Claude Opus 4200K
Claude Sonnet 4200K
Claude 3.5 Haiku200K
Sonar Pro200K
GLM-5.1200K
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.
o3 is a reasoning AI model by OpenAI, released on April 16, 2025. It supports a context window of 200K tokens and can generate up to 100K output tokens.
At $2 per million input tokens and $8 per million output tokens, its blended cost of $5.00/1M tokens places it in the premium pricing tier. Its value score of 18.8 reflects the balance of quality and cost.
Using o3 with Swfte
Access o3 through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try o3 in our AI Playground before integrating.
Deploy a model with Swfte Connect
One gateway, every provider, per-token cost visibility. Swap models without touching your code.