Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
321.1
TFLOPS
JM
akierum
104
3
Follow
0 followers
·
1 following
AI & ML interests
None yet
Recent Activity
replied
to
bytkim
's
post
3 days ago
Qwen3.6-27B-pi-tune v2 is coming plus its sibling 35B-A3B variant I just wanted to share an update on the progress of future releases for Qwen3.6-pi-tuned family models. Both 27B and 35B models are now unified under native think/no-think functionality. The biggest lesson I started to learn after reviewing many suggestions: Fine-tuning for local open-weight agents isn't just about raw coding capability or benchmark numbers. Harness fluency, tool-calling, validation loops, and user-facing behavior matter just as much, sometimes more. That insight and philosophy is exactly what v2 is based on. Although it hasn't even been a month since the original release I wanted to get out the 35B-A3B variant as soon as possible due to popular demand. With the current and upcoming releases of a new class of Agentic LLM's (Fable, GPT5.6, etc) expect v3 to be the best yet. For everyone already running the original: what's it doing well, and what makes you reach for a different model? These suggestions help shape future releases. https://huggingface.co/collections/bytkim/qwen36-pi-tune PS: Heres a sneak peek at the newest version with the new 35B-A3B variant performing a one shot task. 😁
new
activity
3 days ago
peculiar-ragdoll/Dirk-Qwen3.8-27B-GGUF:
Where is the normal Q8 version? to fit into 48GB VRAM with 256k context?
new
activity
8 days ago
zai-org/GLM-5.3-Flash:
WE NEED 35B A3B
View all activity
Organizations
None yet
akierum
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
peculiar-ragdoll/Dirk-Qwen3.8-27B-GGUF
3 days ago
Where is the normal Q8 version? to fit into 48GB VRAM with 256k context?
5
#13 opened 4 days ago by
akierum
New activity in
zai-org/GLM-5.3-Flash
8 days ago
WE NEED 35B A3B
👍
🔥
18
3
#3 opened 9 days ago by
AsThirtyThree
New activity in
bytkim/Qwen3.6-27B-MTP-pi-reasoning-GGUF
8 days ago
will there be qwen 3.8 27B version ?
3
#5 opened 12 days ago by
akierum
New activity in
beyoru/KAT-Coder-V2.5-Dev-VL
9 days ago
the bartowski quantization works fine with mmproj from qwen no need for modifications
#1 opened 9 days ago by
akierum
New activity in
Kwaipilot/KAT-Coder-V2.5-Dev
10 days ago
WOW! Not kidding, this model nuts! 💥
❤️
8
4
#29 opened 14 days ago by
auf1r2
New activity in
z-lab/Qwen3.8-27B-DFlash2
10 days ago
for 3090 MTP and gguf is best for performance 250k context and accuracy.
1
#8 opened 10 days ago by
akierum
New activity in
DavidAU/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-NM-DAU-NEO-MAX-MTP-GGUF
11 days ago
Recommended Sampling, etc Settings
🧠
1
2
#18 opened 13 days ago by
Shad3
New activity in
unsloth/Qwen3.8-27B-GGUF
12 days ago
Why Qwen3.8-27B-UD-Q8_K_XL.gguf is not updated like other quants?
7
#91 opened 15 days ago by
akierum
New activity in
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
12 days ago
why Huihui-Qwen3.8-27B-abliterated-Q8_0_L.gguf is bigger than unsloth Q8-XL?
1
#7 opened 12 days ago by
akierum
New activity in
unsloth/Qwen3.8-27B-GGUF
13 days ago
Chat template issue and ways to solve it
❤️
👍
7
6
#42 opened 20 days ago by
Snol-std
What's the use of mtp-Qwen3.8-27B-Q4_0.gguf?
7
#97 opened 14 days ago by
cnayan01
New activity in
DavidAU/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-NM-DAU-NEO-MAX-MTP-GGUF
14 days ago
request (424318 tokens) exceeds the available context size (250112 tokens), try increasing it: {"error":{"code":400,"message":"request (424318 tokens)
2
#14 opened 15 days ago by
akierum
New activity in
unsloth/Qwen3.8-27B-GGUF
15 days ago
Introducing Unsloth Dynamic v3 Qwen3.8
👍
🔥
52
35
#74 opened 16 days ago by
danielhanchen
New activity in
redashes/Qwen3.8-27B-BF16-SSMFIX
16 days ago
can you make gguf?
6
#1 opened 18 days ago by
Gavin-chen
New activity in
Qwen/Qwen3.8-27B
18 days ago
DFlash draft model instead of MTP one
👍
2
9
#54 opened 21 days ago by
artden111
It is inferior to the qwen3.6 35 model in image and video recognition
4
#69 opened 20 days ago by
wzgrx
Why Qwen3.8-27B overthinks? Here the reason and partial fix confirmed by benchmarks.
❤️
👀
25
29
#76 opened 20 days ago by
LuffyTheFox
New activity in
unsloth/Qwen3.8-27B-GGUF
20 days ago
codeneedle benchmark results
2
#33 opened 20 days ago by
akierum
New activity in
unsloth/Qwen3.8-27B-GGUF
21 days ago
Performance report on RTX 5090: 100 t/s with UD-Q6_K_XL
10
#14 opened 21 days ago by
SlavikF
New activity in
Qwen/Qwen3.8-2.4T-A95B
21 days ago
建议延续马云的思想将千问团队剥离出来独立发展,借助资本做大做强
👍
1
1
#23 opened 22 days ago by
emei8
Load more