SheepNav
新上线今天0 投票

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models

arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised fine-tuning (SFT) to pursue the same underlying goals. When projects share a goal, we should test whether lessons learned from one area transfer to the other areas. We study three such transfers, each taking a lesson developed in one SFT setting and testing it in another. First, we port a lesson abo

延伸阅读

  1. 梯度与自然梯度之间:LoRA 初始化的连续统
  2. 弱到强在线蒸馏:让弱小模型教会强大模型的新范式
  3. 动态参数化不等于动态推理:新研究揭示AI模型效率评估误区
查看原文