whichllm — Browse and compare AI model specs and pricing

Cortecs

llama-3.1-nemotron-ultra-253b-v1 model ID, context window & pricing

models.dev synced record

Quick facts

Model ID llama-3.1-nemotron-ultra-253b-v1
Source Cortecs
Context Window 128000
Pricing $0.60 input / $1.79 output per 1M tokens
Capabilities tool calling, reasoning, structured output

Model overview

llama-3.1-nemotron-ultra-253b-v1 is an AI model from Cortecs with 128000 token context window and text input support.

Published pricing is $0.60 input and $1.79 output per 1M tokens.

  • Workloads that use text inputs with text outputs.
  • Agent and tool workflows that need function calling.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID llama-3.1-nemotron-ultra-253b-v1
Provider Cortecs
Family -
Status -
Knowledge Cutoff -
Release Date 2025-04-07
Input Modalities text
Output Modalities text
Context Window 128000
Input Limit -
Output Limit 128000
Tool Calling Yes
Reasoning Yes
Structured Output Yes
Temperature Control No
Open Weights No
Input Cost / 1M tokens $0.60
Output Cost / 1M tokens $1.79
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -