LLM reasoning / chain-of-thought
single-teacher CoT distillation
Superseded baseline#94 of 772 most-superseded
Cited as a baseline — critiqued by newer work, not yet beaten on a benchmark here
2 papers critique it · 0 beat it on benchmarks
What papers say
Verbatim critique sentences, each from a paper that cites single-teacher CoT distillation as a baseline.
Traditional approaches force student SLMs to fit a single teacher-generated rationale, neglecting the diversity in reasoning paths that could lead to the SLM simply mimics the superficial style of the teacher model's outputs
“single-teacher optimization may introduce bias”
What to use instead
Recent methods in the same sub-problem, not yet superseded in the knowledge base — arXiv benchmark leaders, not vetted production recommendations.
- Oct 15, 2025