Skip to content
Not available in this workspace
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for nvidia

NVIDIA: Llama 3.1 Nemotron 70B Instruct

nvidia/llama-3.1-nemotron-70b-instruct

Model weights

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforcement Learning from Human Feedback (RLHF), it excels in automatic alignment benchmarks. This model is tailored for applications requiring high accuracy in helpfulness and response generation, suitable for diverse user queries across multiple domains.

Usage of this model is subject to Meta's Acceptable Use Policy(opens in new tab).

Modalities

Context

131K

Released

Oct 15, 2024

Knowledge Cutoff

Dec 2023

ActivityFAQ

Activity

Token volume and request traffic to this model over time.

About NVIDIA: Llama 3.1 Nemotron 70B Instruct

OpenRouter makes NVIDIA: Llama 3.1 Nemotron 70B Instruct available through a unified, OpenAI-compatible API using the model ID nvidia/llama-3.1-nemotron-70b-instruct.

NVIDIA: Llama 3.1 Nemotron 70B Instruct accepts text and returns text. It has a 131,072-token context window.

It was released on October 15, 2024; its knowledge cutoff is December 31, 2023.

More models from Nvidia

  • Nemotron 3.5 ASR Streaming Multilingual 0.6B
  • Nemotron 3.5 Lightning
  • Nemotron 3 Embed 1B (free)

Frequently asked questions

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforcement Learning from Human Feedback (RLHF), it excels in automatic alignment benchmarks.

Llama 3.1 Nemotron 70B Instruct has a 131,072 token context window.

Nemotron 3.5 Lightning, Nemotron 3.5 Content Safety (free), Nemotron 3 Ultra and 5 more are other text models from Nvidia.

Llama 3.1 Nemotron 70B Instruct was released on October 15, 2024. Its knowledge cutoff is December 31, 2023.