Tag
This paper analyzes how prompt template selection during knowledge distillation affects the safety alignment of student large language models, finding that chat templates lead to greater degradation compared to non-chat templates across multiple models and benchmarks.
A retrospective on a multi-agent content pipeline that underperformed compared to a simple prompt and template, analyzing the specific failure points.