whichllm — Browse and compare AI model specs and pricing

evroc

Llama-3.3-70B-Instruct model ID, context window & pricing

llama

Quick facts

Model ID nvidia/Llama-3.3-70B-Instruct-FP8
Source evroc
Context Window 128000
Pricing $1.15 input / $1.15 output per 1M tokens
Capabilities tool calling, temperature control, open weights

Model overview

Llama-3.3-70B-Instruct is an AI model from evroc with 128000 token context window and text input support.

Published pricing is $1.15 input and $1.15 output per 1M tokens.

  • Workloads that use text inputs with text outputs.
  • Agent and tool workflows that need function calling.
Model ID nvidia/Llama-3.3-70B-Instruct-FP8
Provider evroc
Family llama
Status -
Knowledge Cutoff 2023-12
Release Date 2024-12-06
Input Modalities text
Output Modalities text
Context Window 128000
Input Limit -
Output Limit 4096
Tool Calling Yes
Reasoning No
Structured Output -
Temperature Control Yes
Open Weights Yes
Input Cost / 1M tokens $1.15
Output Cost / 1M tokens $1.15
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -