Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...
The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Google: Gemini 2.5 Progoogle/gemini-2.5-pro | 32.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GLM-4.6zhipuai/glm-4.6 | 29.5 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 26.4 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| GLM-4.5zhipuai/glm-4.5 | 26.3 | 131.072K | $0.6 | $2.2 | 2025-07-28 | |||
| OpenAI: GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 24.2 | 128K | $5 | $15 | 2024-05-13 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 23.8 | 131.072K | $0.2 | $1.1 | 2025-07-28 | |||
| Devstral 2mistral/devstral-2512 | 23.7 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Mistral Large 3mistral/mistral-large-2512 | 22.7 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | 22.2 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| OpenAI: GPT-4 Turboopenai/gpt-4-turbo | 21.5 | 128K | $10 | $30 | 2023-11-06 | |||
| OpenAI: GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 16.7 | 128K | $2.5 | $10 | 2024-11-20 | |||
| OpenAI: GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 16.6 | 128K | $2.5 | $10 | 2024-08-06 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 15.9 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 13.8 | 131.072K | $2 | $6 | 2024-11-18 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 13.6 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 10.7 | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| Google: Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite | 9.5 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |