Datasets
← Home
5 datasets found (tag “code”)
Filter by tag: sft (17) · reasoning (12) · distilled (9) · agentic (8) · tool-calling (8) · code (5) · instruction-tuning (5) · web-text (4) · pretraining (4) · math (4) · filtered (3) · multi-turn (3) · multilingual (3) · open-harness (3) · deepseek-r1 (2) · reinforcement-learning (2) · dpo (2) · deepseek-v4 (2) · preference (2) · deepseek-v3.2 (2) · synthetic (2) · function-calling (2) · conversations (2) · human-written (2) · opencode (1) · openhands (1) · qwen3-coder (1) · real-user (1) · research (1) · rlhf (1)
Sort: popular · downloads · stars · newest
- Nemotron-SFT-OpenCode-v1 — Agentic instruction-tuning data for the OpenCode CLI framework. (1 downloads, 1 stars, cc-by-4.0 (per Nemotron post-training family; not independently re-verified on this specific card), tool-calling, agentic, skills)
- Nemotron-SWE-v1 — 59k agent trajectories targeting SWE-Bench-style repo navigation and issue-fixing. (0 downloads, 0 stars, cc-by-4.0 (a few subsets are bsd-3-clause — see HF dataset viewer), tool-calling, agentic, code)
- qwen3-coder-480b-distill-mini — 9,543 cleaned code-reasoning samples distilled from Qwen3-Coder-480B-A35B-Instruct. (1 downloads, 3 stars, apache-2.0, code, reasoning, distilled)
- Deepseek-v4-pro-max-distill-1500x — Coding and math reasoning traces distilled from DeepSeek V4 Pro Max. (0 downloads, 0 stars, unspecified (source: DeepSeek V4 Pro Max, MIT + explicit distillation permission), reasoning, distilled, code)
- Llama-Nemotron-Post-Training-Dataset — NVIDIA's 30M+ example SFT+RL post-training dataset behind Llama-3-Nemotron. (0 downloads, 0 stars, cc-by-4.0, sft, reinforcement-learning, reasoning)