Will this get the DFlash + DDTree acceleration treatment?

#4
by TomLucidor - opened

If this model as an agent can beat SOTA, it should be able to use SpecDec to make things faster, right? https://huggingface.co/z-lab/Qwen3.5-35B-A3B-DFlash
P.S. Will 4-bit quants and Ternary LM be supported in the near future?

Intern Science org

@BoZhang for clarification, are the 4B model suitable for accelerating Agents-A1? It's not Eagle3 nor DFlash, so there has to be a way to make output faster

Sign up or log in to comment