All Models
llama-3.1-nemotron-ultra-253b-v1
A reasoning-optimized LLM based on Llama 3.1, Nemotron Ultra 253B delivers strong performance in tasks like RAG and tool use, with high efficiency and reduced latency.
Available Providers (2)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
| | llama-3.1-nemotron-ultra-253b-v1 | $0.60/MTok | $1.79/MTok | 128K | 128K | |
| | nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 | $0.60/MTok | $1.80/MTok | 128K | 4.1K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output