openai/gpt-5.1-codex
Codex GPT for repository edits, code review, and practical software agents
- Context
- 400K
- Input
- $1.07/M
- Output
- $8.5/M
Filter 978 source-linked models by creator, price, context, modality, and published benchmark coverage.
openai/gpt-5.1-codex
Codex GPT for repository edits, code review, and practical software agents
openai/gpt-5.1-codex-mini
Coding-optimized GPT model for repository edits, reviews, and agentic software work
openai/gpt-5.1-chat-latest
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
moonshotai/kimi-k2-thinking
Thinking Kimi model for slower research passes, planning, and hard technical questions
moonshotai/kimi-k2-thinking-turbo
Kimi reasoning model for long-horizon research, planning, and tool use
anthropic/claude-opus-4-5-20251101
Flagship Claude model for deep reasoning, coding, and long-horizon agents
openai/gpt-oss-safeguard-120b
Safety model for policy screening, moderation, and risk-aware routing workflows
nvidia/llama-3.1-nemotron-safety-guard-8b-v3
Safety model for policy screening, moderation, and risk-aware routing workflows
nvidia/nemotron-nano-12b-v2-vl
Nemotron multimodal model for visual reasoning and agentic AI workflows
minimax/MiniMax-M2
Efficient open MiniMax model built for coding agents and tool-heavy workflows
anthropic/claude-haiku-4-5-20251001
Fast Claude model for responsive assistance, classification, and lightweight agents
anthropic/claude-haiku-4-5
Fast Claude lane for lightweight agents, office tasks, and responsive chat
google/veo-3.1-generate-preview
Video model for prompt-guided generation, editing, and motion workflows
google/veo-3.1-fast-generate-preview
Video model for prompt-guided generation, editing, and motion workflows
google/gemini-2.5-computer-use-preview-10-2025
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks
openai/gpt-5-pro
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
zhipuai/glm-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
google/gemini-2.5-pro-tts
Speech generation model for controllable voice, narration, and audio delivery
google/gemini-2.5-flash-tts
Speech generation model for controllable voice, narration, and audio delivery
anthropic/claude-sonnet-4-5
Balanced Claude model for coding, analysis, agent workflows, and cost control
anthropic/claude-sonnet-4-5-20250929
Balanced Claude model for coding, analysis, agent workflows, and cost control
alibaba/qwen3-max
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
alibaba/qwen3-vl-plus
Qwen vision-language model for visual reasoning, documents, and agent tasks
openai/gpt-5-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
sarvam/sarvam-105b
Flagship Indian-language reasoning model for enterprise multilingual applications
alibaba/qwen3-next-80b-a3b-instruct
Qwen instruction model for multilingual chat, reasoning, and tool use
alibaba/qwen3-next-80b-a3b-thinking
Efficient Qwen thinking model for local reasoning, math, and coding agents
cohere/command-a-translate-08-2025
Translation model for multilingual conversion, localization, and cross-language workflows
google/gemini-2.5-flash-image
Nano Banana image model for fast generation, edits, and character-consistent assets
cohere/command-a-reasoning-08-2025
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
nvidia/nemotron-nano-9b-v2
Compact Nemotron model for efficient reasoning and deployable AI agents
zhipuai/glm-4.5v
GLM vision model for visual reasoning, documents, and multimodal agents
openai/gpt-5
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
openai/gpt-5-mini
Small GPT-5 for responsive agents, coding help, and everyday automation
openai/gpt-5-nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
openai/gpt-5-chat-latest
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
openai/gpt-oss-120b
Open GPT reasoning model for self-hosted agents and controllable deployments
openai/gpt-oss-20b
Open GPT reasoning model for self-hosted agents and controllable deployments
anthropic/claude-opus-4-1-20250805
Flagship Claude model for deep reasoning, coding, and long-horizon agents
anthropic/claude-opus-4-1
Flagship Claude model for deep reasoning, coding, and long-horizon agents
cohere/command-a-vision-07-2025
Cohere vision model for multilingual document analysis, OCR, and image understanding
zhipuai/glm-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
zhipuai/glm-4.5-air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
zhipuai/glm-4.5-flash
Efficient GLM model for fast reasoning, coding, and agent workflows
alibaba/qwen-flash
Efficient Qwen model for fast chat, extraction, and high-volume workloads
alibaba/qwen3-coder-flash
Qwen coding model for software agents, repository edits, and code reasoning
nvidia/llama-3.3-nemotron-super-49b-v1.5
Nemotron model for efficient reasoning, coding, and specialized AI agents
alibaba/qwen3-coder-plus
Hosted Qwen coder for software agents, repo edits, and long-context code
mistral/devstral-small-2507
Mistral coding agent model for repository tasks and software engineering workflows
mistral/devstral-medium-2507
Mistral coding agent model for repository tasks and software engineering workflows
| Model | Creator | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 / $8.5 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 / $1.8 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 / $10 | 2025-11-13 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.6 / $2.5 | 2025-11-06 | |||
| Kimi K2 Thinking Turbomoonshotai/kimi-k2-thinking-turbo | 262.144K | $1.15 / $8 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 / $25 | 2025-11-01 | |||
| GPT OSS Safeguard 120Bopenai/gpt-oss-safeguard-120b | 131.072K | $0.15 / $0.6 | 2025-10-29 | |||
| Llama 3.1 Nemotron Safety Guard 8B v3nvidia/llama-3.1-nemotron-safety-guard-8b-v3 | 128K | - / - | 2025-10-28 | |||
| Nemotron Nano 12B v2 VLnvidia/nemotron-nano-12b-v2-vl | 128K | $0.2 / $0.6 | 2025-10-28 | |||
| MiniMax-M2minimax/MiniMax-M2 | 196.608K | $0.3 / $1.2 | 2025-10-27 | |||
| Claude Haiku 4.5anthropic/claude-haiku-4-5-20251001 | 200K | $1 / $5 | 2025-10-15 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 / $5 | 2025-10-15 | |||
| Veo 3.1 Previewgoogle/veo-3.1-generate-preview | 1.024K | - / - | 2025-10-15 | |||
| Veo 3.1 Fast Previewgoogle/veo-3.1-fast-generate-preview | 1.024K | - / - | 2025-10-15 | |||
| Gemini 2.5 Computer Use Previewgoogle/gemini-2.5-computer-use-preview-10-2025 | 128K | - / - | 2025-10-07 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 / $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 / $2.2 | 2025-09-30 | |||
| Gemini 2.5 Pro TTSgoogle/gemini-2.5-pro-tts | 32.768K | $1 / $20 | 2025-09-30 | |||
| Gemini 2.5 Flash TTSgoogle/gemini-2.5-flash-tts | 32.768K | $0.5 / $10 | 2025-09-30 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 / $15 | 2025-09-29 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 / $15 | 2025-09-29 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 / $6 | 2025-09-23 | |||
| Qwen3-VL Plusalibaba/qwen3-vl-plus | 262.144K | $0.2 / $1.6 | 2025-09-23 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 / $9 | 2025-09-15 | |||
| Sarvam 105Bsarvam/sarvam-105b | 131.072K | $0.04 / $0.16 | 2025-09-01 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 / $2 | 2025-09 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 / $6 | 2025-09 | |||
| Command A Translatecohere/command-a-translate-08-2025 | 8K | $2.5 / $10 | 2025-08-28 | |||
| Nano Bananagoogle/gemini-2.5-flash-image | 32.768K | $0.3 / $30 | 2025-08-26 | |||
| Command A Reasoningcohere/command-a-reasoning-08-2025 | 256K | $2.5 / $10 | 2025-08-21 | |||
| Nemotron Nano 9B v2nvidia/nemotron-nano-9b-v2 | 131.072K | $0.04 / $0.16 | 2025-08-18 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 64K | $0.6 / $1.8 | 2025-08-11 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 / $10 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 / $2 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 / $0.4 | 2025-08-07 | |||
| GPT-5 Chat (latest)openai/gpt-5-chat-latest | 400K | $1.25 / $10 | 2025-08-07 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 / $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.03 / $0.13 | 2025-08-05 | |||
| Claude Opus 4.1anthropic/claude-opus-4-1-20250805 | 200K | $15 / $75 | 2025-08-05 | |||
| Claude Opus 4.1 (latest)anthropic/claude-opus-4-1 | 200K | $15 / $75 | 2025-08-05 | |||
| Command A Visioncohere/command-a-vision-07-2025 | 128K | $2.5 / $10 | 2025-07-31 | |||
| GLM-4.5zhipuai/glm-4.5 | 131.072K | $0.6 / $2.2 | 2025-07-28 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 131.072K | $0.2 / $1.1 | 2025-07-28 | |||
| GLM-4.5-Flashzhipuai/glm-4.5-flash | 131.072K | - / - | 2025-07-28 | |||
| Qwen Flashalibaba/qwen-flash | 1M | $0.05 / $0.4 | 2025-07-28 | |||
| Qwen3 Coder Flashalibaba/qwen3-coder-flash | 1M | $0.3 / $1.5 | 2025-07-28 | |||
| Llama 3.3 Nemotron Super 49B v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.1 / $0.4 | 2025-07-25 | |||
| Qwen3 Coder Plusalibaba/qwen3-coder-plus | 1.04858M | $1 / $5 | 2025-07-23 | |||
| Devstral Smallmistral/devstral-small-2507 | 128K | $0.1 / $0.3 | 2025-07-10 | |||
| Devstral Mediummistral/devstral-medium-2507 | 128K | $0.4 / $2 | 2025-07-10 |