OpenRouter AI Model Pricing - September 2026
📝 Markdown RenderedComplete OpenRouter pricing reference for Nemotron 3 Ultra, GLM 5.2, and GLM 5.3 Flash with provider comparisons, benchmarks, and performance data.
OpenRouter AI Model Pricing รขยย September 2026
Compiled from openrouter.ai on September 9, 2026.
NVIDIA: Nemotron 3 Ultra (free)
| Field | Value |
|---|---|
| Slug | nvidia/nemotron-3-ultra-550b-a55b:free |
| Price | Free |
| Context | 1,000,000 tokens |
| Max Output | 65,536 tokens |
| Released | June 4, 2026 |
| Architecture | 550B total / 55B active parameters (MoE), hybrid Transformer-Mamba |
| Modalities | Text in / Text out |
| Tool Calling | Yes (tools tool_choice) |
| Structured Output | No (response_format not supported) |
| Best For | Long-running agentic workflows, agent orchestration, coding agents, deep research, complex enterprise tasks |
| Performance | ~4 tok/s throughput, ~38.8s latency (P50), 96.44% uptime |
| Top Apps | Hermes Agent, Kilo Code, Claude Code, Janitor AI, OpenClaw |
Comments
No comments yet. Start the discussion.