27 results Clear filters
Ranked by raw score for Artificial Analysis Coding Index - . Versions are kept separate, and incompatible results are never combined.
GPT-5-Codexby OpenAI
2025-09-15
openai/gpt-5-codex

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$1.1/M
Output
$9/M
Step 3.7 Flashby StepFun
2026-05-29
stepfun/step-3.7-flash

Newer StepFun flash model for faster agents, coding, and multimodal prompts

T Reasoning Tools Open weights
Context
256K
Input
$0.185/M
Output
$1.11/M
2026-04-02
stepfun/step-3.5-flash-2603

StepFun flash model for efficient multimodal reasoning, coding, and tool use

T Reasoning Tools Open weights
Context
256K
Input
$0.1/M
Output
$0.3/M
2026-06-09
cohere/north-mini-code-1-0

Cohere coding model for practical software engineering and agentic edits

T Reasoning Tools Structured output Open weights
Context
256K
Input
-
Output
-
2025-06-17
google/gemini-2.5-pro

Google's proven reasoning model for coding, math, and multimodal analysis

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.25/M
Output
$10/M
Step 3.5 Flashby StepFun
2026-01-29
stepfun/step-3.5-flash

StepFun flash lane for quick multimodal reasoning and coding assistance

T Reasoning Tools Open weights
Context
256K
Input
$0.1/M
Output
$0.3/M
GLM-4.6by Zhipu AI
2025-09-30
zhipuai/glm-4.6

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

T Reasoning Tools Open weights
Context
204.8K
Input
$0.6/M
Output
$2.2/M
Qwen3 Maxby Alibaba Qwen
2025-09-23
alibaba/qwen3-max

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

T Tools Weight access not listed
Context
262.144K
Input
$1.2/M
Output
$6/M
GLM-4.5by Zhipu AI
2025-07-28
zhipuai/glm-4.5

Hybrid-reasoning GLM release that made the 4.5 line broadly useful

T Reasoning Tools Open weights
Context
131.072K
Input
$0.6/M
Output
$2.2/M
Mistral Small 4by Mistral AI
2026-03-16
mistral/mistral-small-2603

Fast Mistral production model for chat, extraction, and cost-sensitive agents

T Reasoning Tools Open weights
Context
256K
Input
$0.15/M
Output
$0.6/M
2024-05-13
openai/gpt-4o-2024-05-13

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Tools Structured output Weight access not listed
Context
128K
Input
$5/M
Output
$15/M
GLM-4.5-Airby Zhipu AI
2025-07-28
zhipuai/glm-4.5-air

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

T Reasoning Tools Open weights
Context
131.072K
Input
$0.2/M
Output
$1.1/M
Devstral 2by Mistral AI
2025-12-09
mistral/devstral-2512

Mistral's coding-agent model for repository work, terminal tasks, and software fixes

T Tools Open weights
Context
262.144K
Input
$0.4/M
Output
$2/M
Mistral Large 3by Mistral AI
2024-11-01
mistral/mistral-large-2512

Mistral's largest general model for enterprise agents, coding, and multilingual reasoning

T Tools Open weights
Context
262.144K
Input
$0.5/M
Output
$1.5/M
2025-06-17
google/gemini-2.5-flash

Fast Gemini workhorse for multimodal apps where latency and price matter

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.3/M
Output
$2.5/M
GPT-4 Turboby OpenAI
2023-11-06
openai/gpt-4-turbo

Compact GPT model for low-latency assistance and high-volume workloads

T Tools Weight access not listed
Context
128K
Input
$10/M
Output
$30/M
2025-04
alibaba/qwen3-coder-30b-a3b-instruct

Smaller Qwen coder for efficient local agents and repo-level fixes

T Tools Open weights
Context
262.144K
Input
$0.45/M
Output
$2.25/M
2024-11-20
openai/gpt-4o-2024-11-20

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
2024-08-06
openai/gpt-4o-2024-08-06

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
DeepSeek-R1by DeepSeek
2025-01-20
deepseek/deepseek-r1

Classic open reasoning model for transparent math, coding, and deliberate problem solving

T Reasoning Tools Open weights
Context
128K
Input
$0.7/M
Output
$2.5/M
Mistral Large 2.1by Mistral AI
2024-11-18
mistral/mistral-large-2411

Flagship Mistral model for advanced reasoning, coding, and multilingual work

T Tools Open weights
Context
131.072K
Input
$2/M
Output
$6/M
Mistral Medium 3by Mistral AI
2025-05-07
mistral/mistral-medium-2505

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
GPT-4by OpenAI
2023-11-06
openai/gpt-4

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Tools Weight access not listed
Context
8.192K
Input
$30/M
Output
$60/M
GLM-4.5Vby Zhipu AI
2025-08-11
zhipuai/glm-4.5v

GLM vision model for visual reasoning, documents, and multimodal agents

T Reasoning Tools Open weights
Context
64K
Input
$0.6/M
Output
$1.8/M
2024-12-06
meta/llama-3.3-70b-instruct

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

T Tools Open weights
Context
128K
Input
$0.13/M
Output
$0.4/M
GPT-3.5-turboby OpenAI
2023-03-01
openai/gpt-3.5-turbo

Compact GPT model for low-latency assistance and high-volume workloads

T Tools Weight access not listed
Context
16.385K
Input
$0.5/M
Output
$1.5/M
2025-06-17
google/gemini-2.5-flash-lite

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.1/M
Output
$0.4/M