AIAny
Icon for item

Qwen3.8-Max Distillation 50K

A curated collection of 49,772 teacher-generated chat traces from qwen3.8-max-preview for supervised fine-tuning and off-policy distillation. Preserves visible chain-of-thought blocks, emphasizes math/code/reasoning mixes, and includes provenance and licensing cautions tied to Alibaba Cloud Model Studio.

Introduction

The dataset provides long, teacher-produced assistant traces from a large reasoning-capable model (qwen3.8-max-preview), making it useful for experiments in supervised fine-tuning and off-policy distillation that aim to capture chain-of-thought style reasoning. Its mixture is heavily weighted toward math and code tasks and preserves raw model outputs, including visible <think> blocks, but comes with nontrivial provenance and license constraints inherited from upstream sources and Alibaba Cloud Model Studio.

Information

  • Websitehuggingface.co
  • Organizationsr0b0tlab, Alibaba Cloud, Qwen (Alibaba)
  • Published date2026/07/22

Categories

More Items

Hugging Face

Provides live Codex-CLI agent run traces from GPT-5.6 Sol capturing coding, debugging, security reviews, and harness/seed workflows in cumulative next-action prefixes — suitable for supervised fine-tuning and analysis of tool-using coding agents.

Hugging Face

Provides behavior-preserving next-step training traces from Claude Fable 5 for supervised fine-tuning and analysis of instruction-following, tool-calling, and coding agents. Runtime-normalized, independently verified, and supplied as Parquet/JSONL with 13,357 cumulative rows from 2,443 accepted trajectories.

Hugging Face

50,000 distilled conversational traces (≈120M tokens) generated from GLM-5.2 for high-reasoning text generation and QA, covering STEM, programming, creative and support dialogues; Apache-2.0 licensed.