117 results Clear filters
Ranked by raw score for Design Arena: 3d - . Versions are kept separate, and incompatible results are never combined.
Undated
qwen/qwen3-235b-a22b-thinking-2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

T Weight access not listed
Context
131.072K
Input
$0.23/M
Output
$2.3/M
Undated
qwen/qwen3-235b-a22b-2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

T Weight access not listed
Context
262.144K
Input
$0.09/M
Output
$0.55/M
Undated
mistralai/ministral-14b-2512

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

T Weight access not listed
Context
262.144K
Input
$0.2/M
Output
$0.2/M
2025-11-13
openai/gpt-5.1-codex-mini

Coding-optimized GPT model for repository edits, reviews, and agentic software work

T Reasoning Tools Weight access not listed
Context
400K
Input
$0.22/M
Output
$1.8/M
Undated
inception/mercury-2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

T Weight access not listed
Context
128K
Input
$0.25/M
Output
$0.75/M
GPT-5 Nanoby OpenAI
2025-08-07
openai/gpt-5-nano

Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs

T Reasoning Tools Structured output Weight access not listed
Context
400K
Input
$0.05/M
Output
$0.4/M
Undated
openai/gpt-5-nano:batch

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

T Weight access not listed
Context
400K
Input
$0.025/M
Output
$0.2/M
Undated
mistralai/ministral-3b-2512

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

T Weight access not listed
Context
131.072K
Input
$0.1/M
Output
$0.1/M
GPT-4.1 nanoby OpenAI
2025-04-14
openai/gpt-4.1-nano

Tiny GPT-4.1 option for classification, routing, and very high-volume tasks

T Tools Weight access not listed
Context
1.04758M
Input
$0.1/M
Output
$0.4/M
Undated
openai/gpt-oss-120b:free

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

T Weight access not listed
Context
131.072K
Input
Free
Output
Free
GPT OSS 120Bby OpenAI
2025-08-05
openai/gpt-oss-120b

Open GPT reasoning model for self-hosted agents and controllable deployments

T Reasoning Tools Open weights
Context
131.072K
Input
$0.03/M
Output
$0.17/M
Undated
meta-llama/llama-4-maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

T Weight access not listed
Context
1.04858M
Input
$0.2/M
Output
$0.8/M
GPT-4oby OpenAI
2024-05-13
openai/gpt-4o

Omni-era GPT for multimodal chat, practical coding, and general assistants

T Tools Weight access not listed
Context
128K
Input
$2.5/M
Output
$10/M
Qwen: Qwen3 235B A22Bby Alibaba Qwen
Undated
qwen/qwen3-235b-a22b

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

T Weight access not listed
Context
131.072K
Input
$0.455/M
Output
$1.82/M
o4-miniby OpenAI
2025-04-16
openai/o4-mini

Fast o-series model for compact reasoning, coding, and tool use

T Reasoning Tools Structured output Weight access not listed
Context
200K
Input
$1.1/M
Output
$4.4/M
GPT-4.1by OpenAI
2025-04-14
openai/gpt-4.1

Long-lived GPT workhorse for coding, instruction following, and production apps

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$2/M
Output
$8/M
GPT-4.1 miniby OpenAI
2025-04-14
openai/gpt-4.1-mini

Affordable GPT-4.1 lane for fast coding help and structured extraction

T Tools Structured output Weight access not listed
Context
1.04758M
Input
$0.4/M
Output
$1.6/M