No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| moonshotai/Kimi-K2-Thinkingbaseten/moonshotai/kimi-k2-thinking | Not documented | $0.6 | $2.5 | — | |||
| moonshotai/Kimi-K2-Instruct-0905baseten/moonshotai/kimi-k2-instruct-0905 | Not documented | $0.6 | $2.5 | — | |||
| google/gemma-4-E2B-it-qat-mobile-transformersgoogle/gemma-4-E2B-it-qat-mobile-transformers | Not documented | — | — | — | |||
| google/gemma-4-E4B-it-qat-mobile-transformersgoogle/gemma-4-E4B-it-qat-mobile-transformers | Not documented | — | — | — | |||
| minimax/minimax-m2.7novita/minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| meta-llama/CodeLlama-7b-Instruct-hfmeta-llama/CodeLlama-7b-Instruct-hf | Not documented | — | — | — | |||
| meta-llama/CodeLlama-7b-Python-hfmeta-llama/CodeLlama-7b-Python-hf | Not documented | — | — | — | |||
| us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| meta-llama/CodeLlama-7b-hfmeta-llama/CodeLlama-7b-hf | Not documented | — | — | — | |||
| google/gemma-4-E2B-it-qat-mobile-ctgoogle/gemma-4-E2B-it-qat-mobile-ct | Not documented | — | — | — | |||
| google/gemma-4-E4B-it-qat-mobile-ctgoogle/gemma-4-E4B-it-qat-mobile-ct | Not documented | — | — | — | |||
| google/gemma-4-31B-it-qat-w4a16-ctgoogle/gemma-4-31B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| google/gemma-4-12B-it-qat-w4a16-ctgoogle/gemma-4-12B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| openai/gpt-oss-120bbaseten/openai/gpt-oss-120b | Not documented | $0.1 | $0.5 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1.04858M | $0.625 | $5 | — | |||
| chatgpt-4o-latestopenai/chatgpt-4o-latest | 128K | $5 | $15 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| Qwen/Qwen3.8-27BQwen/Qwen3.8-27B | Not documented | — | — | — | |||
| global.anthropic.claude-opus-4-7bedrock_converse/global.anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| us.anthropic.claude-opus-4-7bedrock_converse/us.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| xiaomimimo/mimo-v2.5novita/xiaomimimo/mimo-v2.5 | 1.04858M | $0.168 | $0.336 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| NousResearch/Hermes-4-70Bnebius/nousresearch/hermes-4-70b | 131.072K | $0.13 | $0.4 | — | |||
| google/gemma-4-E4B-it-qat-w4a16-ctgoogle/gemma-4-E4B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| us-gov-east-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| eu.anthropic.claude-opus-4-7bedrock_converse/eu.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| NousResearch/Hermes-4-405Bnebius/nousresearch/hermes-4-405b | 131.072K | $1 | $3 | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-16000microsoft/Dayhoff-170M-UR90-HL-16000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-8000microsoft/Dayhoff-170M-UR90-HL-8000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GR-1000microsoft/Dayhoff-170M-GR-1000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-4000microsoft/Dayhoff-170M-UR90-HL-4000 | Not documented | — | — | — | |||
| google/gemma-4-E2B-it-qat-w4a16-ctgoogle/gemma-4-E2B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| deepseek-ai/DeepSeek-V3.1baseten/deepseek-ai/deepseek-v3.1 | Not documented | $0.5 | $1.5 | — | |||
| deepseek-ai/DeepSeek-V3-0324baseten/deepseek-ai/deepseek-v3-0324 | Not documented | $0.77 | $0.77 | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct-FP8gmi/qwen/qwen3-vl-235b-a22b-instruct-fp8 | 262.144K | $0.3 | $1.4 | — | |||
| zai-org/GLM-4.7-FP8gmi/zai-org/glm-4.7-fp8 | 202.752K | $0.4 | $2 | — | |||
| google.gemma-3-12b-itbedrock_converse/google.gemma-3-12b-it | 128K | $0.09 | $0.29 | — | |||
| au.anthropic.claude-opus-4-7bedrock_converse/au.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| moonshotai/Kimi-K3nebius/moonshotai/kimi-k3 | 1.024M | $3 | $15 | — | |||
| google/gemma-4-12Bgoogle/gemma-4-12B | Not documented | — | — | — | |||
| llama3.2-1bsnowflake/llama3.2-1b | 128K | — | — | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| moonshotai/Kimi-K2.7-Codenebius/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — |