ninfer-5080/Qwen3.8-27B-RTX5080
Image-Text-to-Text • Updated • 4.49k • 7
High-performance local inference, NVIDIA RTX 5080 optimisation, CUDA, quantisation, long-context inference, multimodal models, speculative decoding, and reproducible model builds.