The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

meta-llama/llama-3.3-70b-instruct:free 65.536K context Free input Free output

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

z-ai/glm-5.2:free 256K context Free input Free output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output

No provider description is available for this model yet.

zai/glm-4.6 200K context $0.6/M input $2.2/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3:batch 1M context $1/M input $2/M output

GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.

openai/gpt-4o-search-preview 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

cloudflare/@cf/openai/gpt-oss-120b 128K context $0.35/M input $0.75/M output

No provider description is available for this model yet.

cerebras/qwen-3.8-27b 65.536K context $0.99/M input $1.49/M output

No provider description is available for this model yet.

openai/computer-use-preview 8.192K context $3/M input $12/M output

No provider description is available for this model yet.

bedrock/anthropic.claude-v1 100K context $8/M input $24/M output

No provider description is available for this model yet.

bedrock/anthropic.claude-v2:1 100K context $8/M input $24/M output

No provider description is available for this model yet.

deepseek/deepseek-flash 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

anyscale/google/gemma-7b-it 8.192K context $0.15/M input $0.15/M output

No provider description is available for this model yet.

google/medgemma-1.5-4b-it Not documented context Input not listed Output not listed

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

z-ai/glm-4.7-flash 131.072K context $0.061/M input $0.4/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output