Llama-Nemotron-Post-Training-Dataset

← All datasets

Author: Distillio · License: cc-by-4.0 · Format: jsonl

0 downloads · 0 stars · 0 mirrors · size unknown · 1 file

NVIDIA's 30M+ example SFT+RL post-training dataset behind Llama-3-Nemotron.

Tags: sft · reinforcement-learning · reasoning · math · code

Description

NVIDIA's 30M+ example post-training dataset (SFT + RL) spanning math, code, science, chat, and safety, used to train the Llama-3-Nemotron model family.

Download

Direct download: https://huggingface.co/datasets/nvidia/Llama-Nemotron-Post-Training-Dataset/resolve/main/SFT/math/math_v1.1.jsonl

Or via the API: curl -L -o file "https://huggingface.co/datasets/nvidia/Llama-Nemotron-Post-Training-Dataset/resolve/main/SFT/math/math_v1.1.jsonl"

Added 2026-08-11. Platform: Distillio — the open source training data collective.