DistillDetect Reproduction (arXiv:2607.09692)
Unofficial reproduction of the 32 distilled students of arXiv:2607.09692, trained with the authors' released code and data.
-
4B • Updated • 16 -
francescortu/DistillDetect-gemma-3-4b-pt-from-gpt-oss-120b-s1
4B • Updated • 23 -
francescortu/DistillDetect-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-OMI-1K
4B • Updated • 19 -
francescortu/DistillDetect-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-OMI-918-COT
4B • Updated • 26 -
francescortu/DistillDetect-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-s1
4B • Updated • 19 -
francescortu/DistillDetect-gemma-3-4b-pt-from-Qwen3-8B-OMI-1K
4B • Updated • 19 -
francescortu/DistillDetect-gemma-3-4b-pt-from-Qwen3-8B-s1
4B • Updated • 19 -
francescortu/DistillDetect-gemma-3-4b-pt-from-o1-s1
4B • Updated • 17 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-gpt-oss-120b-OMI-1K
3B • Updated • 29 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-gpt-oss-120b-s1
3B • Updated • 26 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-Llama-3.3-70B-Instruct-OMI-1K
3B • Updated • 26 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-Llama-3.3-70B-Instruct-OMI-918-COT
3B • Updated • 27 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-Llama-3.3-70B-Instruct-s1
3B • Updated • 26 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-Qwen3-8B-OMI-1K
3B • Updated • 32 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-Qwen3-8B-s1
3B • Updated • 26 -
francescortu/DistillDetect-Llama-3.2-3B-Instruct-from-o1-s1
3B • Updated • 25 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-gpt-oss-120b-OMI-1K
2B • Updated • 35 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-gpt-oss-120b-s1
2B • Updated • 25 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-OMI-1K
2B • Updated • 34 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-OMI-918-COT
2B • Updated • 20 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-s1
2B • Updated • 23 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-Qwen3-8B-OMI-1K
2B • Updated • 41 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-Qwen3-8B-s1
2B • Updated • 25 -
francescortu/DistillDetect-Qwen2.5-1.5B-from-o1-s1
2B • Updated • 28 -
francescortu/DistillDetect-Qwen2.5-3B-from-gpt-oss-120b-OMI-1K
3B • Updated • 29 -
francescortu/DistillDetect-Qwen2.5-3B-from-gpt-oss-120b-s1
3B • Updated • 20 -
francescortu/DistillDetect-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-OMI-1K
3B • Updated • 24 -
francescortu/DistillDetect-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-OMI-918-COT
3B • Updated • 17 -
francescortu/DistillDetect-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-s1
3B • Updated • 23 -
francescortu/DistillDetect-Qwen2.5-3B-from-Qwen3-8B-OMI-1K
3B • Updated • 17 -
francescortu/DistillDetect-Qwen2.5-3B-from-Qwen3-8B-s1
3B • Updated • 13 -
francescortu/DistillDetect-Qwen2.5-3B-from-o1-s1
3B • Updated • 16
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-gpt-oss-120b-OMI-1K
Text Generation • UpdatedNote Gemma-3-4B-PT <- GPT-OSS-120B / OMI(1K): 13 checkpoints across training. GSM8K 36.9 -> 51.0 (+14.1, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-gpt-oss-120b-s1
Text Generation • UpdatedNote Gemma-3-4B-PT <- GPT-OSS-120B / S1: 13 checkpoints across training. GSM8K 36.9 -> 51.1 (+14.2, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-OMI-1K
Text Generation • UpdatedNote Gemma-3-4B-PT <- Nvidia-Llama-3.3-70B-Instruct / OMI(1K): 13 checkpoints across training. GSM8K 36.9 -> 49.3 (+12.4, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-OMI-918-COT
Text Generation • UpdatedNote Gemma-3-4B-PT <- Nvidia-Llama-3.3-70B-Instruct / OMI(918): 13 checkpoints across training. GSM8K 36.9 -> 52.4 (+15.5, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-Llama-3.3-70B-Instruct-s1
Text Generation • UpdatedNote Gemma-3-4B-PT <- Nvidia-Llama-3.3-70B-Instruct / S1: 13 checkpoints across training. GSM8K 36.9 -> 46.9 (+9.9, p=7e-12)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-Qwen3-8B-OMI-1K
Text Generation • UpdatedNote Gemma-3-4B-PT <- Qwen-3-8B / OMI(1K): 13 checkpoints across training. GSM8K 36.9 -> 52.4 (+15.5, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-Qwen3-8B-s1
Text Generation • UpdatedNote Gemma-3-4B-PT <- Qwen-3-8B / S1: 13 checkpoints across training. GSM8K 36.9 -> 50.8 (+13.9, p=0)
francescortu/DistillDetect-traj-gemma-3-4b-pt-from-o1-s1
Text Generation • UpdatedNote Gemma-3-4B-PT <- o1 / S1: 13 checkpoints across training. GSM8K 36.9 -> 46.8 (+9.9, p=7.2e-11)
francescortu/DistillDetect-traj-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-OMI-1K
Text Generation • UpdatedNote Qwen-2.5-1.5B <- Nvidia-Llama-3.3-70B-Instruct / OMI(1K): 13 checkpoints across training. GSM8K 65.9 -> 68.7 (+2.8, p=0.0074)
francescortu/DistillDetect-traj-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-OMI-918-COT
Text Generation • UpdatedNote Qwen-2.5-1.5B <- Nvidia-Llama-3.3-70B-Instruct / OMI(918): 13 checkpoints across training. GSM8K 65.9 -> 68.6 (+2.7, p=0.015)
francescortu/DistillDetect-traj-Qwen2.5-1.5B-from-Llama-3.3-70B-Instruct-s1
Text Generation • UpdatedNote Qwen-2.5-1.5B <- Nvidia-Llama-3.3-70B-Instruct / S1: 13 checkpoints across training. GSM8K 65.9 -> 69.5 (+3.6, p=0.00025)
francescortu/DistillDetect-traj-Qwen2.5-1.5B-from-Qwen3-8B-s1
Text Generation • UpdatedNote Qwen-2.5-1.5B <- Qwen-3-8B / S1: 13 checkpoints across training. GSM8K 65.9 -> 68.3 (+2.4, p=0.05)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-gpt-oss-120b-s1
Text Generation • UpdatedNote Qwen-2.5-3B <- GPT-OSS-120B / S1: 13 checkpoints across training. GSM8K 75.9 -> 81.3 (+5.5, p=3.1e-06)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-OMI-1K
Text Generation • UpdatedNote Qwen-2.5-3B <- Nvidia-Llama-3.3-70B-Instruct / OMI(1K): 13 checkpoints across training. GSM8K 75.9 -> 79.8 (+3.9, p=0.00015)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-OMI-918-COT
Text Generation • UpdatedNote Qwen-2.5-3B <- Nvidia-Llama-3.3-70B-Instruct / OMI(918): 13 checkpoints across training. GSM8K 75.9 -> 79.1 (+3.2, p=0.0017)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-Llama-3.3-70B-Instruct-s1
Text Generation • UpdatedNote Qwen-2.5-3B <- Nvidia-Llama-3.3-70B-Instruct / S1: 13 checkpoints across training. GSM8K 75.9 -> 79.3 (+3.4, p=0.00094)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-Qwen3-8B-OMI-1K
Text Generation • UpdatedNote Qwen-2.5-3B <- Qwen-3-8B / OMI(1K): 13 checkpoints across training. GSM8K 75.9 -> 81.1 (+5.2, p=1.2e-05)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-Qwen3-8B-s1
Text Generation • UpdatedNote Qwen-2.5-3B <- Qwen-3-8B / S1: 13 checkpoints across training. GSM8K 75.9 -> 80.3 (+4.4, p=0.0002)
francescortu/DistillDetect-traj-Qwen2.5-3B-from-o1-s1
Text Generation • UpdatedNote Qwen-2.5-3B <- o1 / S1: 13 checkpoints across training. GSM8K 75.9 -> 78.7 (+2.8, p=0.023)
francescortu/DistillDetect-normalized-traces
Viewer • Updated • 7.92k • 65Note Teacher traces rewritten to a single ... + answer format, so output format no longer identifies the teacher