Author: Distillio · License: cc-by-4.0 · Format: jsonl
0 downloads · 0 stars · 0 mirrors · size unknown · 1 file
NVIDIA's 30M+ example SFT+RL post-training dataset behind Llama-3-Nemotron.
Tags: sft · reinforcement-learning · reasoning · math · code
NVIDIA's 30M+ example post-training dataset (SFT + RL) spanning math, code, science, chat, and safety, used to train the Llama-3-Nemotron model family.
Direct download: https://huggingface.co/datasets/nvidia/Llama-Nemotron-Post-Training-Dataset/resolve/main/SFT/math/math_v1.1.jsonl
Or via the API: curl -L -o file "https://huggingface.co/datasets/nvidia/Llama-Nemotron-Post-Training-Dataset/resolve/main/SFT/math/math_v1.1.jsonl"
Added 2026-08-11. Platform: Distillio — the open source training data collective.