3,199 models

No provider description is available for this model yet.

snowflake/llama3.3-70b 128K context $0.72/M input $0.72/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1:batch 200K context $7.5/M input $30/M output

No provider description is available for this model yet.

snowflake/llama3.2-3b 128K context Input not listed Output not listed

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

aion-labs/aion-2.0 131.072K context $0.8/M input $1.6/M output

No provider description is available for this model yet.

bedrock/ai21.jamba-1-5-mini-v1:0 256K context $0.2/M input $0.4/M output

No provider description is available for this model yet.

bedrock/ai21.jamba-instruct-v1:0 70K context $0.5/M input $0.7/M output

This model always redirects to the latest Grok model from xAI.

~x-ai/grok-latest 500K context $2/M input $6/M output

This model always redirects to the latest model in the GLM Flash family.

~z-ai/glm-flash-latest 1.04858M context $0.075/M input $0.25/M output

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

moonshotai/kimi-k3:batch 1.04858M context $3/M input $15/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

This model always redirects to the latest model in the GPT Sol family.

~openai/gpt-sol-latest 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

bedrock/eu-north-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

This model always redirects to the latest model in the GPT Astra family.

~openai/gpt-astra-latest 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-08-06 64K context Input not listed Output not listed

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

z-ai/glm-4.7-flash 131.072K context $0.061/M input $0.4/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-05-13 64K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

github_copilot/gpt-4o 64K context Input not listed Output not listed

Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...

inclusionai/ling-3.0-tiny:free 262.144K context Free input Free output

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...

bytedance/ui-tars-1.5-7b 128K context $0.1/M input $0.2/M output

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

qwen/qwen3.5-35b-a3b 256K context $0.312/M input $1.25/M output

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

meta/muse-spark-1.2-contributor 1.04858M context $0.1/M input $0.2/M output

No provider description is available for this model yet.

mistral/labs-leanstral-1-5 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

anthropic/claude-mythos-preview 1M context $10/M input $50/M output

No provider description is available for this model yet.

xai/grok-4.20-multi-agent-0309 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

bedrock/amazon.titan-text-lite-v1 42K context $0.3/M input $0.4/M output

No provider description is available for this model yet.

gmi/deepseek-ai/deepseek-v3.2 163.84K context $0.28/M input $0.4/M output

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8:batch 1M context $2.5/M input $12.5/M output