#weak-to-strong
- Automated Weak-to-Strong Researcher: AI Agents Outperform Humans on Alignment Research Anthropic research
- Weak-to-Strong Generalization via Direct On-Policy Distillation ByteDance / Tsinghua University research
- Weak-to-Strong On-Policy Distillation Microsoft Research / University of Maryland / MBZUAI research
- OPRD: eliciting weak-to-strong generalization with on-policy reverse distillation KAIST AI research