Llama-3.1-Nemotron-70B
by NVIDIA
· NEMOTRON family · released Oct 1, 2024
Open weights
Active
Llama tuned by NVIDIA for helpfulness.
This model has no first-party API. Open-weight release: self-host from published weights.
Specifications
| Model type | Text LLM |
|---|---|
| Architecture | Dense |
| Parameters | 70 |
| Context window | 131.1K tokens |
| Max output | Not disclosed tokens |
| Knowledge cutoff | Apr 1, 2024 |
| License | NVIDIA Open Model License |
| Modalities | Text input Text output |
| Tags | Long context |
| End of life | Not disclosed |
| Last verified | Aug 20, 2026 |