Training & Alignment
2026
6
- Can Closed-Source Models Be Distilled? Knowledge Distillation for Generative Language Models
- Reward Hacking: When Optimizers Reverse-Search the Reward Signal
- The Evolution of Reward Design: From RLHF to RLVR
- Reinforcement Learning in LLM Alignment: From Reward Signals to Advantage Estimation
- Parameter-Efficient Fine-Tuning (PEFT): From Adapter to LoRA
- The Essence of LLM Reasoning and Training: From Surrogates to Reinforcement Learning Geometry
2025
4