4,213 models
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-13b-hf Not documented context Input not listed Output not listed

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

nousresearch/hermes-3-llama-3.1-405b:free 131.072K context Free input Free output

No provider description is available for this model yet.

gmi/deepseek-ai/deepseek-v3-0324 163.84K context $0.28/M input $0.88/M output

No provider description is available for this model yet.

gmi/google/gemini-3-pro-preview 1.04858M context $2/M input $12/M output

No provider description is available for this model yet.

gmi/google/gemini-3-flash-preview 1.04858M context $0.5/M input $3/M output

No provider description is available for this model yet.

gmi/moonshotai/kimi-k2-thinking 262.144K context $0.8/M input $1.2/M output

No provider description is available for this model yet.

gmi/minimaxai/minimax-m2.1 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

baseten/minimaxai/minimax-m2.5 Not documented context $0.3/M input $1.2/M output

No provider description is available for this model yet.

baseten/nvidia/nemotron-120b-a12b Not documented context $0.3/M input $0.75/M output

No provider description is available for this model yet.

baseten/zai-org/glm-5 Not documented context $0.95/M input $3.15/M output

No provider description is available for this model yet.

baseten/zai-org/glm-4.7 Not documented context $0.6/M input $2.2/M output

No provider description is available for this model yet.

baseten/zai-org/glm-4.6 Not documented context $0.6/M input $2.2/M output

No provider description is available for this model yet.

baseten/moonshotai/kimi-k2.5 Not documented context $0.6/M input $3/M output

No provider description is available for this model yet.

baseten/moonshotai/kimi-k2-thinking Not documented context $0.6/M input $2.5/M output

Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...

arcee-ai/coder-large 32.768K context $0.5/M input $0.8/M output
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-7b-Instruct-hf Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-7b-Python-hf Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

meta-llama/CodeLlama-7b-hf Not documented context Input not listed Output not listed

No provider description is available for this model yet.

baseten/openai/gpt-oss-120b Not documented context $0.1/M input $0.5/M output

No provider description is available for this model yet.

openai/chatgpt-4o-latest 128K context $5/M input $15/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.8-27B Not documented context Input not listed Output not listed

This model always redirects to the latest model in the GLM Flash family.

~z-ai/glm-flash-latest 1.04858M context $0.075/M input $0.25/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3.1 Not documented context $0.5/M input $1.5/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3-0324 Not documented context $0.77/M input $0.77/M output