All Models

llama-3.1-nemotron-ultra-253b-v1

Reasoning Tool Calling Open Weights Structured Output

A reasoning-optimized LLM based on Llama 3.1, Nemotron Ultra 253B delivers strong performance in tasks like RAG and tool use, with high efficiency and reduced latency.

Providers 2
Released Jan 15, 2025
Input Modalities text
Output Modalities text
Tarsk Use coding

Available Providers (2)

Provider Model ID Input Cost Output Cost Context Max Output Docs
Cortecs llama-3.1-nemotron-ultra-253b-v1 $0.60/MTok $1.79/MTok 128K 128K
Nebius Token Factory nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 $0.60/MTok $1.80/MTok 128K 4.1K

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output