RunInfra
Inference provider
- Hosted model listings
- 3
- Models created
- 0
- API base
https://api.runinfra.ai/v1- Last verified
- Aug 17, 2026
Inference catalog
Models served by RunInfra
Provider-specific identifiers, limits, and source-linked prices.
| Model | Provider model ID | Context | Input / 1M | Output / 1M | Cached / 1M | Capabilities | Verified |
|---|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0731DeepSeek | deepseek-ai/DeepSeek-V4-Flash-0731 |
1.04858M | $0.13 | $0.27 | $0.01 | ReasoningToolsJSON | Aug 17, 2026 |
| Nemotron 3.5 Lightning 30B A3BNVIDIA | nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 |
262.144K | $0.05 | $0.15 | - | ReasoningToolsJSON | Aug 17, 2026 |
| Qwen3.8 27BAlibaba Qwen | Qwen/Qwen3.8-27B |
262.144K | $0.1 | $0.4 | $0.01 | ReasoningToolsJSON | Aug 17, 2026 |