Checkpoints of OPSA on different base models
Yi Ding
Tuwhy
AI & ML interests
None yet
Recent Activity
authored a paper about 7 hours ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement updated a model about 12 hours ago
Tuwhy/Qwen3-4B-OPSA updated a model about 12 hours ago
Tuwhy/Qwen3.5-9B-OPSA