Text Standard
Price-optimized, high-throughput text
0
Pick the right intelligence tier for your task, or route across them for peak performance at lower costs.
Price-optimized, high-throughput text
Reliable models for everyday tasks
Frontier models for complex reasoning
Price-optimized, high-throughput code
Reliable models for daily coding work
Frontier models for complex coding tasks
Price-optimized, high-throughput agents
Reliable models for daily agent work
Frontier models for complex agent tasks
Latest qualifying route for GPT frontier supply
Latest qualifying route for Claude Opus frontier supply
Latest qualifying route for Gemini Pro frontier supply
Latest qualifying route for MiniMax supply
Latest qualifying route for GLM supply
Latest qualifying route for DeepSeek Pro supply
Latest qualifying route for Kimi supply
Latest qualifying route for ByteDance Pro supply
1Prices are refreshed frequently, but they move in real time with the market. For the latest price, visit The Grid App.
2The qualifying model list changes as new models meet the instrument spec and existing models are updated. Any model and supplier that qualifies the performance specification can serve requests on that instrument. The models listed here were recently active but are not guaranteed at any given time.
3Lab Latest launch prices are single blended rates calculated from the current platform token mix: 10.30% input, 84.07% cache read, 3.39% cache write, and 2.25% output/reasoning. The mix is based on production usage from May 25–June 18, 2026.
Point each task to the right instrument to see improvements in performance and cost efficiency.
Classify requests on Standard first. Send onlythose that need reasoning to Prime.
Run retrieval and reranking on Standard. Generatethe final output on Prime.
Keep your production workload on Prime. Use Maxonly for large inputs or where accuracy is critical.
Learn more on our blog