494 results Clear filters
2025-10-07
google/gemini-2.5-computer-use-preview-10-2025

Specialized Gemini 2.5 model for browser-control agents that automate UI tasks

T Weight access not listed
Context
128K
Input
-
Output
-
GPT-5 Proby OpenAI
2025-10-06
openai/gpt-5-pro

Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning

T Reasoning Tools Weight access not listed
Context
400K
Input
$15/M
Output
$120/M
GLM-4.6by Zhipu AI
2025-09-30
zhipuai/glm-4.6

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

T Reasoning Tools Open weights
Context
204.8K
Input
$0.6/M
Output
$2.2/M
2025-09-29
anthropic/claude-sonnet-4-5

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Reasoning Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Claude Sonnet 4.5by Anthropic
2025-09-29
anthropic/claude-sonnet-4-5-20250929

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Reasoning Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Qwen3 Maxby Alibaba Qwen
2025-09-23
alibaba/qwen3-max

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

T Tools Weight access not listed
Context
262.144K
Input
$1.2/M
Output
$6/M
Qwen3-VL Plusby Alibaba Qwen
2025-09-23
alibaba/qwen3-vl-plus

Qwen vision-language model for visual reasoning, documents, and agent tasks

T Reasoning Tools Weight access not listed
Context
262.144K
Input
$0.2/M
Output
$1.6/M
GPT-5-Codexby OpenAI
2025-09-15
openai/gpt-5-codex

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$1.1/M
Output
$9/M
Sarvam 105Bby Sarvam
2025-09-01
sarvam/sarvam-105b

Flagship Indian-language reasoning model for enterprise multilingual applications

T Reasoning Tools Open weights
Context
131.072K
Input
$0.04/M
Output
$0.16/M
2025-09
alibaba/qwen3-next-80b-a3b-instruct

Qwen instruction model for multilingual chat, reasoning, and tool use

T Tools Open weights
Context
131.072K
Input
$0.5/M
Output
$2/M
2025-09
alibaba/qwen3-next-80b-a3b-thinking

Efficient Qwen thinking model for local reasoning, math, and coding agents

T Reasoning Tools Open weights
Context
131.072K
Input
$0.5/M
Output
$6/M
2025-08-21
cohere/command-a-reasoning-08-2025

Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows

T Reasoning Tools Open weights
Context
256K
Input
$2.5/M
Output
$10/M
2025-08-18
nvidia/nemotron-nano-9b-v2

Compact Nemotron model for efficient reasoning and deployable AI agents

T Reasoning Tools Open weights
Context
131.072K
Input
$0.04/M
Output
$0.16/M
GPT-5by OpenAI
2025-08-07
openai/gpt-5

Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$1.25/M
Output
$10/M
GPT-5 Miniby OpenAI
2025-08-07
openai/gpt-5-mini

Small GPT-5 for responsive agents, coding help, and everyday automation

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.25/M
Output
$2/M
GPT-5 Nanoby OpenAI
2025-08-07
openai/gpt-5-nano

Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.05/M
Output
$0.4/M
2025-08-07
openai/gpt-5-chat-latest

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

T Reasoning Structured output Weight access not listed
Context
400K
Input
$1.25/M
Output
$10/M
GPT OSS 120Bby OpenAI
2025-08-05
openai/gpt-oss-120b

Open GPT reasoning model for self-hosted agents and controllable deployments

T Reasoning Tools Open weights
Context
131.072K
Input
$0.03/M
Output
$0.17/M
GPT OSS 20Bby OpenAI
2025-08-05
openai/gpt-oss-20b

Open GPT reasoning model for self-hosted agents and controllable deployments

T Reasoning Tools Open weights
Context
131.072K
Input
$0.03/M
Output
$0.13/M
Claude Opus 4.1by Anthropic
2025-08-05
anthropic/claude-opus-4-1-20250805

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
2025-08-05
anthropic/claude-opus-4-1

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
2025-07-31
cohere/command-a-vision-07-2025

Cohere vision model for multilingual document analysis, OCR, and image understanding

T Open weights
Context
128K
Input
$2.5/M
Output
$10/M
GLM-4.5by Zhipu AI
2025-07-28
zhipuai/glm-4.5

Hybrid-reasoning GLM release that made the 4.5 line broadly useful

T Reasoning Tools Open weights
Context
131.072K
Input
$0.6/M
Output
$2.2/M
GLM-4.5-Airby Zhipu AI
2025-07-28
zhipuai/glm-4.5-air

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

T Reasoning Tools Open weights
Context
131.072K
Input
$0.2/M
Output
$1.1/M
GLM-4.5-Flashby Zhipu AI
2025-07-28
zhipuai/glm-4.5-flash

Efficient GLM model for fast reasoning, coding, and agent workflows

T Reasoning Tools Weight access not listed
Context
131.072K
Input
-
Output
-
Qwen Flashby Alibaba Qwen
2025-07-28
alibaba/qwen-flash

Efficient Qwen model for fast chat, extraction, and high-volume workloads

T Reasoning Tools Weight access not listed
Context
1M
Input
$0.05/M
Output
$0.4/M
Qwen3 Coder Flashby Alibaba Qwen
2025-07-28
alibaba/qwen3-coder-flash

Qwen coding model for software agents, repository edits, and code reasoning

T Tools Weight access not listed
Context
1M
Input
$0.3/M
Output
$1.5/M
2025-07-25
nvidia/llama-3.3-nemotron-super-49b-v1.5

Nemotron model for efficient reasoning, coding, and specialized AI agents

T Reasoning Tools Open weights
Context
131.072K
Input
$0.1/M
Output
$0.4/M
Qwen3 Coder Plusby Alibaba Qwen
2025-07-23
alibaba/qwen3-coder-plus

Hosted Qwen coder for software agents, repo edits, and long-context code

T Tools Weight access not listed
Context
1.04858M
Input
$1/M
Output
$5/M
Devstral Smallby Mistral AI
2025-07-10
mistral/devstral-small-2507

Mistral coding agent model for repository tasks and software engineering workflows

T Tools Open weights
Context
128K
Input
$0.1/M
Output
$0.3/M
Devstral Mediumby Mistral AI
2025-07-10
mistral/devstral-medium-2507

Mistral coding agent model for repository tasks and software engineering workflows

T Tools Weight access not listed
Context
128K
Input
$0.4/M
Output
$2/M
Mistral Small 3.2by Mistral AI
2025-06-20
mistral/mistral-small-2506

Efficient Mistral model for fast chat, extraction, and production assistants

T Tools Open weights
Context
128K
Input
$0.1/M
Output
$0.3/M
2025-06-17
google/gemini-2.5-flash-lite

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.1/M
Output
$0.4/M
2025-06-17
google/gemini-2.5-flash

Fast Gemini workhorse for multimodal apps where latency and price matter

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$0.3/M
Output
$2.5/M
2025-06-17
google/gemini-2.5-pro

Google's proven reasoning model for coding, math, and multimodal analysis

T Reasoning Tools Structured output Weight access not listed
Context
1.04858M
Input
$1.25/M
Output
$10/M
2025-06-11
nvidia/mistral-nemotron

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

T Open weights
Context
128K
Input
-
Output
-
o3-proby OpenAI
2025-06-10
openai/o3-pro

High-effort o3 tier for difficult technical reasoning and careful answers

T Reasoning Tools Weight access not listed
Context
200K
Input
$20/M
Output
$80/M
2025-05-22
anthropic/claude-opus-4-0

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
Claude Opus 4by Anthropic
2025-05-22
anthropic/claude-opus-4-20250514

Flagship Claude model for deep reasoning, coding, and long-horizon agents

T Reasoning Tools Weight access not listed
Context
200K
Input
$15/M
Output
$75/M
Claude Sonnet 4by Anthropic
2025-05-22
anthropic/claude-sonnet-4-20250514

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Reasoning Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
2025-05-22
anthropic/claude-sonnet-4-0

Balanced Claude model for coding, analysis, agent workflows, and cost control

T Reasoning Tools Weight access not listed
Context
200K
Input
$3/M
Output
$15/M
Mistral Medium 3by Mistral AI
2025-05-07
mistral/mistral-medium-2505

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
o3by OpenAI
2025-04-16
openai/o3

Deliberate o-series reasoner for hard math, coding, and multi-step analysis

T Reasoning Tools Structured output Weight access not listed
Context
200K
Input
$2/M
Output
$8/M
o4-miniby OpenAI
2025-04-16
openai/o4-mini

Fast o-series model for compact reasoning, coding, and tool use

T Reasoning Tools Structured output Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
2025-04-15
nvidia/llama-3.1-nemotron-70b-instruct

Nemotron model for efficient reasoning, coding, and specialized AI agents

T Tools Open weights
Context
128K
Input
-
Output
-
GPT-4.1by OpenAI
2025-04-14
openai/gpt-4.1

Long-lived GPT workhorse for coding, instruction following, and production apps

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$2/M
Output
$8/M
GPT-4.1 miniby OpenAI
2025-04-14
openai/gpt-4.1-mini

Affordable GPT-4.1 lane for fast coding help and structured extraction

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$0.4/M
Output
$1.6/M
GPT-4.1 nanoby OpenAI
2025-04-14
openai/gpt-4.1-nano

Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

T Tools Weight access not listed
Context
1.04758M
Input
$0.1/M
Output
$0.4/M
2025-04-07
nvidia/llama-3.3-nemotron-super-49b-v1

Nemotron model for efficient reasoning, coding, and specialized AI agents

T Reasoning Tools Open weights
Context
131.072K
Input
-
Output
-
2025-04-07
nvidia/llama-3.1-nemotron-ultra-253b

Flagship Nemotron model for high-throughput reasoning and complex agents

T Reasoning Tools Open weights
Context
128K
Input
$0.6/M
Output
$1.8/M