3,646 models

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

cognitivecomputations/dolphin-mistral-24b-venice-edition:free 32.768K context Free input Free output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-fast 1.04858M context $4.5/M input $22.5/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.6 262.144K context $1.045/M input $4.4/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.7-code 262.144K context $1.05/M input $4.4/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-us 1.04858M context $3.3/M input $16.5/M output

No provider description is available for this model yet.

fireworks_ai/qwen3p8-max 262.144K context $2/M input $6/M output

No provider description is available for this model yet.

fireworks_ai/muse-glimmer-30b 131.072K context $0.35/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.5 262.144K context $0.66/M input $3.3/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k3 1.04858M context $3.3/M input $16.5/M output

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

mistralai/mistral-large-2407 131.072K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/fw-minimax-m3 512K context $0.33/M input $1.32/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

fireworks_ai/glm-5p2-fast 1.04858M context $2.1/M input $6.6/M output

No provider description is available for this model yet.

azure_ai/grok-4.3 200K context $1.25/M input $2.5/M output

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...

openai/gpt-3.5-turbo-16k 16.385K context $3/M input $4/M output

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash:batch 1.04858M context $0.75/M input $4.5/M output

No provider description is available for this model yet.

fireworks_ai/deepseek-v4-flash-0731 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

github_copilot/gpt-4.1 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4.1-2025-04-14 128K context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/fw-inkling 1.04858M context $1/M input $4.05/M output

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.

openai/gpt-3.5-turbo-instruct 4.095K context $1.5/M input $2/M output