129 results Clear filters
Ranked by raw score for Agentic Index - . Versions are kept separate, and incompatible results are never combined.
Claude Opus 5by Anthropic
2026-07-24
anthropic/claude-opus-5

Strongest Claude Opus model for coding, agents, and professional work

T Reasoning Tools Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
GPT-5.6 Solby OpenAI
2026-07-09
openai/gpt-5.6-sol

Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$5/M
Output
$30/M
Claude Fable 5by Anthropic
2026-06-09
anthropic/claude-fable-5

Claude model for creative writing, analysis, and controlled agent workflows

T Reasoning Tools Weight access not listed
Context
1M
Input
$10/M
Output
$50/M
Undated
anthropic/claude-fable-5:batch

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

T Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
Kimi K3by Moonshot AI
2026-07-16
moonshotai/kimi-k3

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

T Reasoning Tools Structured output Open weights
Context
1.04858M
Input
$3/M
Output
$15/M
GPT-5.6 Terraby OpenAI
2026-07-09
openai/gpt-5.6-terra

Balanced GPT-5.6 model for capable, cost-efficient everyday work

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$2/M
Output
$12/M
Undated
anthropic/claude-opus-4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

T Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
anthropic/claude-opus-4.8:batch

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

T Weight access not listed
Context
1M
Input
$2.5/M
Output
$12.5/M
Claude Sonnet 5by Anthropic
2026-06-30
anthropic/claude-sonnet-5

Everyday Claude agent model for coding, planning, browsing, and general work

T Reasoning Tools Weight access not listed
Context
1M
Input
$2/M
Output
$10/M
anthropic/claude-sonnet-5:batch

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

T Weight access not listed
Context
1M
Input
$1/M
Output
$5/M
Undated
x-ai/grok-4.5

Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

T Weight access not listed
Context
500K
Input
$2/M
Output
$6/M
GPT-5.6 Lunaby OpenAI
2026-07-09
openai/gpt-5.6-luna

Cost-efficient GPT-5.6 model for fast, high-volume workloads

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$0.2/M
Output
$1.2/M
GPT-5.5by OpenAI
2026-04-23
openai/gpt-5.5

Default frontier GPT for coding, computer use, research, and knowledge work

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$5/M
Output
$30/M
Undated
openai/gpt-5.5:batch

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

T Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
Undated
anthropic/claude-opus-4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

T Weight access not listed
Context
1M
Input
$5/M
Output
$25/M
anthropic/claude-opus-4.7:batch

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

T Weight access not listed
Context
1M
Input
$2.5/M
Output
$12.5/M
Undated
z-ai/glm-5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

T Weight access not listed
Context
1.024M
Input
$1.19/M
Output
$3.74/M
GPT-5.4by OpenAI
2026-03-05
openai/gpt-5.4

Agent-ready GPT for coding and computer-use workflows at a lower cost

T Reasoning Tools Structured output Weight access not listed
Context
1.05M
Input
$2.5/M
Output
$15/M
Undated
openai/gpt-5.4:batch

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

T Weight access not listed
Context
1.05M
Input
$1.25/M
Output
$7.5/M
Undated
anthropic/claude-sonnet-4.6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

T Weight access not listed
Context
1M
Input
$3/M
Output
$15/M
2026-07-21
google/gemini-3.6-flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$7.5/M
google/gemini-3.6-flash:batch

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

T Weight access not listed
Context
1.04858M
Input
$0.75/M
Output
$3.75/M
2026-04-08
meta/muse-spark-1.1

Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

T Reasoning Tools Structured output Weight access not listed
Context
1M
Input
$1.25/M
Output
$4.25/M
2026-05-19
google/gemini-3.5-flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.5/M
Output
$9/M
google/gemini-3.5-flash:batch

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

T Weight access not listed
Context
1.04858M
Input
$0.75/M
Output
$4.5/M
DeepSeek V4 Proby DeepSeek
2026-04-24
deepseek/deepseek-v4-pro

Open MoE flagship with million-token context for coding and long agent runs

T Reasoning Tools Structured output Open weights
Context
1M
Input
$0.435/M
Output
$0.87/M
Undated
minimax/minimax-m3

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

T Weight access not listed
Context
524.288K
Input
$0.3/M
Output
$1.2/M
Undated
minimax/minimax-m3:batch

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

T Weight access not listed
Context
524.288K
Input
$0.15/M
Output
$0.6/M
Inklingby Thinkingmachines
2026-07-15
thinkingmachines/inkling

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

T Reasoning Tools Structured output Open weights
Context
1.04858M
Input
$1.87/M
Output
$4.68/M
2026-04-24
deepseek/deepseek-v4-flash

Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work

T Reasoning Tools Structured output Open weights
Context
1M
Input
$0.14/M
Output
$0.28/M
Undated
nex-agi/nex-n2-pro

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

T Weight access not listed
Context
262.144K
Input
$0.25/M
Output
$1/M
Hy3 previewby Tencent
2026-04-20
tencent/hy3-preview

Tencent Hy reasoning model for coding, instruction following, and agent tasks

T Reasoning Tools Open weights
Context
256K
Input
$0.063/M
Output
$0.21/M
Qwen: Qwen3.7 Maxby Alibaba Qwen
Undated
qwen/qwen3.7-max

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

T Weight access not listed
Context
1M
Input
$1.475/M
Output
$4.425/M
Kimi K2.6by Moonshot AI
2026-04-21
moonshotai/kimi-k2.6

Multimodal Kimi workhorse for agent loops, coding tasks, and visual context

T Reasoning Tools Structured output Open weights
Context
262.144K
Input
$0.95/M
Output
$4/M
GPT-5.4 miniby OpenAI
2026-03-17
openai/gpt-5.4-mini

Strong small GPT for coding subagents, quick tool use, and high-volume work

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.75/M
Output
$4.5/M
Undated
openai/gpt-5.4-mini:batch

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

T Weight access not listed
Context
400K
Input
$0.375/M
Output
$2.25/M
Undated
z-ai/glm-5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

T Weight access not listed
Context
200K
Input
$0.966/M
Output
$3.036/M
Kimi K2.7 Codeby Moonshot AI
2026-06-12
moonshotai/kimi-k2.7-code

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

T Reasoning Tools Open weights
Context
262.144K
Input
$0.95/M
Output
$4/M
MiMo-V2.5-Proby Xiaomi
2026-04-22
xiaomi/mimo-v2.5-pro

Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution

T Reasoning Tools Open weights
Context
1.04858M
Input
$0.435/M
Output
$0.87/M
Undated
x-ai/grok-build-0.1

Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

T Weight access not listed
Context
256K
Input
$1/M
Output
$2/M
Qwen: Qwen3.6 Plusby Alibaba Qwen
Undated
qwen/qwen3.6-plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

T Weight access not listed
Context
1M
Input
$0.325/M
Output
$1.95/M
GPT-5.4 nanoby OpenAI
2026-03-17
openai/gpt-5.4-nano

Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.2/M
Output
$1.25/M
Undated
openai/gpt-5.4-nano:batch

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

T Weight access not listed
Context
400K
Input
$0.1/M
Output
$0.625/M
2026-06-04
nvidia/nemotron-3-ultra-550b-a55b

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

T Reasoning Tools Structured output Open weights
Context
1M
Input
$0.5/M
Output
$2.5/M
nvidia/nemotron-3-ultra-550b-a55b:free

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

T Weight access not listed
Context
1M
Input
Free
Output
Free
Qwen: Qwen3.6 27Bby Alibaba Qwen
Undated
qwen/qwen3.6-27b

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

T Weight access not listed
Context
262.144K
Input
$0.3/M
Output
$2/M
2026-07-21
google/gemini-3.5-flash-lite

Fast Gemini model balancing multimodal reasoning, tool use, and cost

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.3/M
Output
$2.5/M
google/gemini-3.5-flash-lite:batch

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

T Weight access not listed
Context
1.04858M
Input
$0.15/M
Output
$1.25/M
GPT-5by OpenAI
2025-08-07
openai/gpt-5

Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$1.25/M
Output
$10/M
Undated
openai/gpt-5:batch

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

T Weight access not listed
Context
400K
Input
$0.625/M
Output
$5/M