sakana/fugu
Multi-agent model for routing expert agents across complex analytical tasks
- Context
- 1M
- Input
- -
- Output
- -
Filter 30 source-linked models by creator, price, context, modality, and published benchmark coverage.
sakana/fugu
Multi-agent model for routing expert agents across complex analytical tasks
sakana/fugu-ultra
Quality-first multi-agent model for hard research, analysis, and competitions
alibaba/qwen3.7-max
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
google/gemini-2.5-pro
Google's proven reasoning model for coding, math, and multimodal analysis
openai/gpt-5-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
stepfun/step-3.5-flash
StepFun flash lane for quick multimodal reasoning and coding assistance
stepfun/step-3.7-flash
Newer StepFun flash model for faster agents, coding, and multimodal prompts
google/gemini-2.5-flash
Fast Gemini workhorse for multimodal apps where latency and price matter
stepfun/step-3.5-flash-2603
StepFun flash model for efficient multimodal reasoning, coding, and tool use
zhipuai/glm-4.6
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
alibaba/qwen3-max
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
mistral/mistral-small-2603
Fast Mistral production model for chat, extraction, and cost-sensitive agents
mistral/mistral-large-2512
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
deepseek/deepseek-r1
Classic open reasoning model for transparent math, coding, and deliberate problem solving
zhipuai/glm-4.5
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
openai/gpt-4o-2024-11-20
GPT model for general reasoning, writing, coding, and tool-assisted tasks
mistral/devstral-2512
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
mistral/mistral-medium-2505
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
openai/gpt-4o-2024-08-06
GPT model for general reasoning, writing, coding, and tool-assisted tasks
openai/gpt-4-turbo
Compact GPT model for low-latency assistance and high-volume workloads
openai/gpt-4o-2024-05-13
GPT model for general reasoning, writing, coding, and tool-assisted tasks
zhipuai/glm-4.5-air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
mistral/mistral-large-2411
Flagship Mistral model for advanced reasoning, coding, and multilingual work
alibaba/qwen3-coder-30b-a3b-instruct
Smaller Qwen coder for efficient local agents and repo-level fixes
meta/llama-3.3-70b-instruct
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
openai/gpt-4o-mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
perplexity/sonar
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
perplexity/sonar-pro
Deeper Sonar search model with broader retrieval and stronger synthesis
zhipuai/glm-4.5v
GLM vision model for visual reasoning, documents, and multimodal agents
google/gemini-2.5-flash-lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
| Model | Creator | Raw benchmark score | Input types | Context | Input / Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Fugusakana/fugu | 60.1 | 1M | - / - | 2026-06-15 | |||
| Fugu Ultrasakana/fugu-ultra | 58.7 | 1M | $5 / $30 | 2026-06-15 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 53.5 | 1M | $2.5 / $7.5 | 2026-05-21 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 / $10 | 2025-06-17 | |||
| GPT-5-Codexopenai/gpt-5-codex | 40.9 | 400K | $1.1 / $9 | 2025-09-15 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 40.4 | 256K | $0.1 / $0.3 | 2026-01-29 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 40.0 | 256K | $0.185 / $1.11 | 2026-05-29 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 / $2.5 | 2025-06-17 | |||
| Step 3.5 Flash 2603stepfun/step-3.5-flash-2603 | 38.5 | 256K | $0.1 / $0.3 | 2026-04-02 | |||
| GLM-4.6zhipuai/glm-4.6 | 38.4 | 204.8K | $0.6 / $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 / $6 | 2025-09-23 | |||
| Mistral Small 4mistral/mistral-small-2603 | 38.0 | 256K | $0.15 / $0.6 | 2026-03-16 | |||
| Mistral Large 3mistral/mistral-large-2512 | 36.2 | 262.144K | $0.5 / $1.5 | 2024-11-01 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 35.7 | 128K | $0.7 / $2.5 | 2025-01-20 | |||
| GLM-4.5zhipuai/glm-4.5 | 34.8 | 131.072K | $0.6 / $2.2 | 2025-07-28 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 33.3 | 128K | $2.5 / $10 | 2024-11-20 | |||
| Devstral 2mistral/devstral-2512 | 33.1 | 262.144K | $0.4 / $2 | 2025-12-09 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 33.1 | 131.072K | $0.4 / $2 | 2025-05-07 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 33.1 | 128K | $2.5 / $10 | 2024-08-06 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 31.9 | 128K | $10 / $30 | 2023-11-06 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 30.9 | 128K | $5 / $15 | 2024-05-13 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 30.6 | 131.072K | $0.2 / $1.1 | 2025-07-28 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 29.2 | 131.072K | $2 / $6 | 2024-11-18 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 27.8 | 262.144K | $0.45 / $2.25 | 2025-04 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 26.0 | 128K | $0.13 / $0.4 | 2024-12-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 22.9 | 128K | $0.15 / $0.6 | 2024-07-18 | |||
| Sonarperplexity/sonar | 22.9 | 128K | $1 / $1 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 22.6 | 200K | $3 / $15 | 2024-01-01 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 22.1 | 64K | $0.6 / $1.8 | 2025-08-11 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 19.3 | 1.04858M | $0.1 / $0.4 | 2025-06-17 |