128 results Clear filters
Ranked by raw score for Intelligence Index - . Versions are kept separate, and incompatible results are never combined.
Kimi K2 Thinkingby Moonshot AI
2025-11-06
moonshotai/kimi-k2-thinking

Thinking Kimi model for slower research passes, planning, and hard technical questions

T Tools Structured output Open weights
Context
262.144K
Input
$0.6/M
Output
$2.5/M
Undated
qwen/qwen3-next-80b-a3b-thinking

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

T Weight access not listed
Context
131.072K
Input
$0.15/M
Output
$1.2/M
Undated
mistralai/mistral-large-2512

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

T Weight access not listed
Context
262.144K
Input
$0.5/M
Output
$1.5/M
Undated
openai/o3-mini-high

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

T Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
Undated
deepseek/deepseek-chat-v3-0324

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

T Weight access not listed
Context
163.84K
Input
$0.27/M
Output
$1.12/M
GPT OSS 20Bby OpenAI
2025-08-05
openai/gpt-oss-20b

Open GPT reasoning model for self-hosted agents and controllable deployments

T Reasoning Tools Open weights
Context
131.072K
Input
$0.03/M
Output
$0.13/M
Undated
openai/gpt-oss-20b:free

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

T Weight access not listed
Context
131.072K
Input
Free
Output
Free
GPT-4.1 miniby OpenAI
2025-04-14
openai/gpt-4.1-mini

Affordable GPT-4.1 lane for fast coding help and structured extraction

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$0.4/M
Output
$1.6/M
Undated
mistralai/mistral-medium-3.1

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

T Weight access not listed
Context
131.072K
Input
$0.4/M
Output
$2/M
Undated
qwen/qwen3-30b-a3b-thinking-2507

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

T Weight access not listed
Context
81.92K
Input
$0.2/M
Output
$2.4/M
Undated
meta-llama/llama-4-maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

T Weight access not listed
Context
1.04858M
Input
$0.2/M
Output
$0.8/M
2025-12-15
nvidia/nemotron-3-nano-30b-a3b

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

T Reasoning Tools Open weights
Context
262.144K
Input
$0.05/M
Output
$0.2/M
nvidia/nemotron-3-nano-30b-a3b:free

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

T Weight access not listed
Context
256K
Input
Free
Output
Free
Undated
inclusionai/ling-2.6-flash

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

T Weight access not listed
Context
262.144K
Input
$0.01/M
Output
$0.03/M
Undated
upstage/solar-pro-3

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...

T Weight access not listed
Context
128K
Input
$0.15/M
Output
$0.6/M
Qwen: Qwen3 32Bby Alibaba Qwen
Undated
qwen/qwen3-32b

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

T Weight access not listed
Context
40.96K
Input
$0.08/M
Output
$0.28/M
Undated
mistralai/ministral-14b-2512

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

T Weight access not listed
Context
262.144K
Input
$0.2/M
Output
$0.2/M
Qwen: Qwen3 14Bby Alibaba Qwen
Undated
qwen/qwen3-14b

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

T Weight access not listed
Context
131.072K
Input
$0.227/M
Output
$0.91/M
Undated
meta-llama/llama-4-scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

T Weight access not listed
Context
327.68K
Input
$0.1/M
Output
$0.3/M
GPT-4.1 nanoby OpenAI
2025-04-14
openai/gpt-4.1-nano

Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

T Tools Weight access not listed
Context
1.04758M
Input
$0.1/M
Output
$0.4/M
meta-llama/llama-3.3-70b-instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

T Weight access not listed
Context
131.072K
Input
$0.13/M
Output
$0.4/M
meta-llama/llama-3.3-70b-instruct:free

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

T Weight access not listed
Context
65.536K
Input
Free
Output
Free
Undated
mistralai/ministral-8b-2512

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

T Weight access not listed
Context
262.144K
Input
$0.15/M
Output
$0.15/M
Qwen: Qwen3 8Bby Alibaba Qwen
Undated
qwen/qwen3-8b

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

T Weight access not listed
Context
131.072K
Input
$0.117/M
Output
$0.455/M
meta-llama/llama-3.1-8b-instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.08/M
Undated
google/gemma-3-27b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.08/M
Output
$0.45/M
Undated
mistralai/ministral-3b-2512

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

T Weight access not listed
Context
131.072K
Input
$0.1/M
Output
$0.1/M
Undated
google/gemma-3-12b-it

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

T Weight access not listed
Context
131.072K
Input
$0.05/M
Output
$0.15/M