Model: Qwen/Qwen3-Coder-30B-A3B-Instruct Method: prune Compression ratio: 0.25 Dataset: hardcoded Samples: 5 Max seq len: 512 Batch size: 4 Device: cuda Torch dtype: auto