986 results
Undated
openai/gpt-5.4:batch

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

T Weight access not listed
Context
1.05M
Input
$1.25/M
Output
$7.5/M
google/gemini-3.1-pro-preview:batch

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

T Weight access not listed
Context
1.04858M
Input
$1/M
Output
$6/M
anthropic/claude-opus-4.6:batch

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

T Weight access not listed
Context
1M
Input
$2.5/M
Output
$12.5/M
google/gemini-3-flash-preview:batch

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

T Weight access not listed
Context
1.04858M
Input
$0.25/M
Output
$1.5/M
Undated
openai/gpt-5.2:batch

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...

T Weight access not listed
Context
400K
Input
$0.875/M
Output
$7/M
anthropic/claude-opus-4.5:batch

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

T Weight access not listed
Context
200K
Input
$2.5/M
Output
$12.5/M
Undated
openai/gpt-5.1:batch

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...

T Weight access not listed
Context
400K
Input
$0.625/M
Output
$5/M
anthropic/claude-haiku-4.5:batch

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

T Weight access not listed
Context
200K
Input
$0.5/M
Output
$2.5/M
anthropic/claude-sonnet-4.5:batch

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

T Weight access not listed
Context
1M
Input
$1.5/M
Output
$7.5/M
Undated
openai/gpt-5:batch

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

T Weight access not listed
Context
400K
Input
$0.625/M
Output
$5/M
Undated
openai/gpt-5-mini:batch

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

T Weight access not listed
Context
400K
Input
$0.125/M
Output
$1/M
Undated
openai/gpt-5-nano:batch

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

T Weight access not listed
Context
400K
Input
$0.025/M
Output
$0.2/M
anthropic/claude-opus-4.1:batch

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

T Weight access not listed
Context
200K
Input
$7.5/M
Output
$37.5/M
google/gemini-2.5-flash-lite:batch

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

T Weight access not listed
Context
1.04858M
Input
$0.05/M
Output
$0.2/M
google/gemini-2.5-flash:batch

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

T Weight access not listed
Context
1.04858M
Input
$0.15/M
Output
$1.25/M
google/gemini-2.5-pro:batch

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

T Weight access not listed
Context
1.04858M
Input
$0.625/M
Output
$5/M
microsoft/Mage-VLby Microsoft
Undated
microsoft/Mage-VL

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Dayhoff-3b-UR90-30000

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
microsoft/Dayhoff-170M-GRS-SS-134000

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-token-translation-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
Undated
deepseek-ai/ESFT-gate-math-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-token-math-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-token-code-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-token-intent-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-token-summary-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
nvidia/Riva-Translate-4B-Instruct-v2

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
nvidia/Riva-Translate-4B-Instruct-v1.1

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
nvidia/Riva-Translate-4B-Instruct

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek/deepseek-v4-flash-0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

T Weight access not listed
Context
1.04858M
Input
$0.14/M
Output
$0.28/M
Undated
microsoft/Dayhoff-3b-UR90-20350

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Dayhoff-3b-UR90-10000

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
Undated
microsoft/Dayhoff-3b-UR90-10

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
microsoft/Dayhoff-170M-GRS-SS-86000

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
microsoft/Dayhoff-170M-GRS-SS-38000

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-
deepseek-ai/ESFT-gate-intent-lite

No provider description is available for this model yet.

Weight access not listed
Context
Not documented
Input
-
Output
-
deepseek-ai/DeepSeek-V4-Flash-0731

No provider description is available for this model yet.

Open weights
Context
Not documented
Input
-
Output
-