138 results Clear filters
Ranked by raw score for Coding Index - . Versions are kept separate, and incompatible results are never combined.
Undated
qwen/qwen3-235b-a22b-thinking-2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

T Weight access not listed
Context
131.072K
Input
$0.23/M
Output
$2.3/M
GPT-4 Turboby OpenAI
2023-11-06
openai/gpt-4-turbo

Compact GPT model for low-latency assistance and high-volume workloads

T Tools Weight access not listed
Context
128K
Input
$10/M
Output
$30/M
Undated
deepseek/deepseek-chat-v3-0324

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

T Weight access not listed
Context
163.84K
Input
$0.27/M
Output
$1.12/M
Kimi K2 Thinkingby Moonshot AI
2025-11-06
moonshotai/kimi-k2-thinking

Thinking Kimi model for slower research passes, planning, and hard technical questions

T Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$2.5/M
GPT OSS 20Bby OpenAI
2025-08-05
openai/gpt-oss-20b

Open GPT reasoning model for self-hosted agents and controllable deployments

T Reasoning Tools Open weights
Context
131.072K
Input
$0.03/M
Output
$0.13/M
Undated
openai/gpt-oss-20b:free

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

T Weight access not listed
Context
131.072K
Input
Free
Output
Free
Undated
mistralai/mistral-medium-3.1

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
GPT-4.1 miniby OpenAI
2025-04-14
openai/gpt-4.1-mini

Affordable GPT-4.1 lane for fast coding help and structured extraction

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$0.4/M
Output
$1.6/M
Undated
mistralai/mistral-large-2512

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

T Weight access not listed
Context
262.144K
Input
$0.5/M
Output
$1.5/M
Undated
qwen/qwen3-next-80b-a3b-thinking

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

T Weight access not listed
Context
131.072K
Input
$0.15/M
Output
$1.2/M
Undated
meta-llama/llama-4-maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

T Weight access not listed
Context
1.04858M
Input
$0.2/M
Output
$0.8/M
Undated
openai/o3-mini-high

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

T Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
Undated
upstage/solar-pro-3

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...

T Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
GPT-5 Miniby OpenAI
2025-08-07
openai/gpt-5-mini

Small GPT-5 for responsive agents, coding help, and everyday automation

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.25/M
Output
$2/M
Undated
openai/gpt-5-mini:batch

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

T Weight access not listed
Context
400K
Input
$0.125/M
Output
$1/M
Qwen: Qwen3 32Bby Alibaba Qwen
Undated
qwen/qwen3-32b

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

T Weight access not listed
Context
40.96K
Input
$0.08/M
Output
$0.28/M
Undated
mistralai/ministral-14b-2512

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

T Weight access not listed
Context
262.144K
Input
$0.2/M
Output
$0.2/M
2025-12-15
nvidia/nemotron-3-nano-30b-a3b

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

T Reasoning Tools Open weights
Context
262.144K
Input
$0.05/M
Output
$0.2/M
nvidia/nemotron-3-nano-30b-a3b:free

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

T Weight access not listed
Context
256K
Input
Free
Output
Free
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

T Weight access not listed
Context
256K
Input
Free
Output
Free
Qwen: Qwen3 14Bby Alibaba Qwen
Undated
qwen/qwen3-14b

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

T Weight access not listed
Context
131.072K
Input
$0.227/M
Output
$0.91/M
GPT-4by OpenAI
2023-11-06
openai/gpt-4

GPT model for general reasoning, writing, coding, and tool-assisted tasks

T Tools Weight access not listed
Context
8.192K
Input
$30/M
Output
$60/M
Undated
qwen/qwen3-30b-a3b-thinking-2507

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

T Weight access not listed
Context
81.92K
Input
$0.2/M
Output
$2.4/M
meta-llama/llama-3.3-70b-instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

T Weight access not listed
Context
131.072K
Input
$0.13/M
Output
$0.4/M
meta-llama/llama-3.3-70b-instruct:free

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

T Weight access not listed
Context
65.536K
Input
Free
Output
Free
GPT-4o miniby OpenAI
2024-07-18
openai/gpt-4o-mini

Small omni GPT for cheap multimodal assistance and production-scale traffic

T Tools Structured output Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
GPT-4.1 nanoby OpenAI
2025-04-14
openai/gpt-4.1-nano

Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

T Tools Weight access not listed
Context
1.04758M
Input
$0.1/M
Output
$0.4/M
GPT-3.5-turboby OpenAI
2023-03-01
openai/gpt-3.5-turbo

Compact GPT model for low-latency assistance and high-volume workloads

T Tools Weight access not listed
Context
16.385K
Input
$0.5/M
Output
$1.5/M
Undated
google/gemma-3-27b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.08/M
Output
$0.45/M
Undated
mistralai/ministral-8b-2512

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

T Weight access not listed
Context
262.144K
Input
$0.15/M
Output
$0.15/M
Undated
ibm-granite/granite-4.1-8b

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

T Open weights
Context
131.072K
Input
$0.05/M
Output
$0.1/M
Qwen: Qwen3 8Bby Alibaba Qwen
Undated
qwen/qwen3-8b

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

T Weight access not listed
Context
131.072K
Input
$0.117/M
Output
$0.455/M
Undated
meta-llama/llama-4-scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

T Weight access not listed
Context
327.68K
Input
$0.1/M
Output
$0.3/M
Undated
google/gemma-3-12b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.15/M
meta-llama/llama-3.1-8b-instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.08/M
Undated
mistralai/ministral-3b-2512

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

T Weight access not listed
Context
131.072K
Input
$0.1/M
Output
$0.1/M
Undated
google/gemma-3n-e4b-it

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

T Weight access not listed
Context
32.768K
Input
$0.06/M
Output
$0.12/M
Undated
google/gemma-3-4b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.1/M