New discussion

VERISON: AWQ For Vllm?

1
#98 opened 11 days ago by
imesh101

Qwen3.8-27B-Brainwaves

👍 1
#93 opened about 1 month ago by
nightmedia

Qwen Qwen3.8 27B? :)

2
#89 opened about 1 month ago by
jpsequeira

Qwen3.8-27B-Fable-Fusion pls

2
#88 opened about 1 month ago by
jezzza1401

toolcalls leaking?

2
#80 opened about 1 month ago by
Mk2Oracle

Somewhat better quants here

#79 opened about 1 month ago by
NikiKrutan

Unsloth UD-Q5_K_XL vs DavidAU Q5_K_S

2
#75 opened about 1 month ago by
calisti

Stable on Hermes IQ4_NL

🤗👍 1
11
#74 opened about 1 month ago by
cgregd

qwen 3.6-35b-a3b 711?

👍 3
4
#71 opened about 2 months ago by
kingsfightger

Mixture-of-Experts version?

👍 2
2
#67 opened about 2 months ago by
wiselibs

qwen 3.8 27b ?

3
#63 opened about 2 months ago by
GravyLuver

This Is Awesome!

👍 1
1
#61 opened about 2 months ago by
austinjohn934

VERSION: MoQ version

👍🔥 5
5
#59 opened about 2 months ago by
Jianqiao1

INT8 W8A8 Quant for vLLM?

👍 3
1
#51 opened about 2 months ago by
bSun0000

VERSION: DFLASH

#47 opened about 2 months ago by
DavidAU

Best llama.cpp server serve settings

1
#30 opened about 2 months ago by
aspmaker

ROCmPFX

❤️ 5
#29 opened about 2 months ago by
FREAKOJC