Datasets
← Home
5 datasets found (tag “instruction-tuning”)
Filter by tag: sft (17) · reasoning (12) · distilled (9) · agentic (8) · tool-calling (8) · code (5) · instruction-tuning (5) · web-text (4) · pretraining (4) · math (4) · filtered (3) · multi-turn (3) · multilingual (3) · open-harness (3) · deepseek-r1 (2) · reinforcement-learning (2) · dpo (2) · deepseek-v4 (2) · preference (2) · deepseek-v3.2 (2) · synthetic (2) · function-calling (2) · conversations (2) · human-written (2) · opencode (1) · openhands (1) · qwen3-coder (1) · real-user (1) · research (1) · rlhf (1)
Sort: popular · downloads · stars · newest
- Hermes 3 Dataset — NousResearch's official post-training dataset behind the Hermes model line. (0 downloads, 1 stars, apache-2.0, sft, instruction-tuning, reasoning)
- OpenHermes-2.5 — ~1M GPT-4-generated instruction pairs deduplicated to ShareGPT/ChatML format. (0 downloads, 0 stars, unknown, sft, instruction-tuning, gpt4-distilled)
- Tulu 3 SFT Mixture — AI2's ~939k-example multilingual SFT corpus with full provenance labels. (1 downloads, 0 stars, odc-by, sft, instruction-tuning, multilingual)
- SmolTalk — Curated Magpie-Ultra-based instruction mix used to train the SmolLM2 family. (0 downloads, 0 stars, mixed (apache-2.0 + source licenses), sft, instruction-tuning, magpie)
- SmolTalk2 — The 2025 SFT mixture behind SmolLM3: decontaminated multi-task instruction data spanning chat, reasoning, and tool use. (0 downloads, 0 stars, mixed (apache-2.0 + source licenses), sft, instruction-tuning, multi-task)