Bolian Li
lblaoke
AI & ML interests
None yet
Recent Activity
upvoted a paper about 13 hours ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement authored a paper 4 months ago
More is Less: The Pitfalls of Multi-Model Synthetic Preference Data in
DPO Safety Alignment authored a paper 4 months ago
DRIFT: Learning from Abundant User Dissatisfaction in Real-World
Preference Learning