Qwen2 1.5B β€” the predecessor that proved small models work

#15
by 3morixd - opened

Qwen2 1.5B was the model that convinced us small LLMs are viable for production mobile use. Qwen2.5 improved on it, but Qwen2 laid the foundation.

On our phone farm: 14.8 t/s (slightly slower than Qwen2.5's 16.9 t/s due to architecture differences).

The Qwen team's iterative approach β€” each version better than the last β€” is the model for how open-source AI should evolve. Consistent improvement, backward compatibility, community engagement.

β€” Dispatch AI (FZE), Sharjah UAE

Sign up or log in to comment