Existing quants are too heavy 😭

#1
by highpolygonal - opened

Hello, could you please create a version suitable for graphics cards with 12GB of VRAM? Something like GSQ-RCO, for example. I think many people would be very grateful. Thank you for your work!

this is just benchmaxxed ai slop probably stick to base model.

Hello, could you please create a version suitable for graphics cards with 12GB of VRAM? Something like GSQ-RCO, for example. I think many people would be very grateful. Thank you for your work!

https://huggingface.co/NikiKrutan/VeriLoop-E2-MTP-GGUF

this is just benchmaxxed ai slop probably stick to base model.

Probably so, but seems to be good from my first impression. Short focused reasoning, job is done faster. Feels very different from the base.

this is just benchmaxxed ai slop probably stick to base model.

The benchmark results were obtained with they're own harness, i am trying to replicate it as pi extension with general rules they provided on model card and technical list. While testing it i am kinda getting why they managed to obtain so high scores, the workflow has some real benefits and if harness enforces them and model is trained for it i see real results.

this is just benchmaxxed ai slop probably stick to base model.

The benchmark results were obtained with they're own harness, i am trying to replicate it as pi extension with general rules they provided on model card and technical list. While testing it i am kinda getting why they managed to obtain so high scores, the workflow has some real benefits and if harness enforces them and model is trained for it i see real results.

How did you do it? share please

Hello, could you please create a version suitable for graphics cards with 12GB of VRAM? Something like GSQ-RCO, for example. I think many people would be very grateful. Thank you for your work!

https://huggingface.co/rodrigoramosrs/veriloop-coder-e2-nvfp4

Hello, could you please create a version suitable for graphics cards with 12GB of VRAM? Something like GSQ-RCO, for example. I think many people would be very grateful. Thank you for your work!

here you go i did it https://huggingface.co/tahaalam2009/VeriLoop-E2-GSQ-RCO-GGUF

Do not want to open new thread. Battle tested this model in GGUF Q4 today. It is very good. I am thinking to switch from Qwen3.8-Flash-Next to this one.

Thank you for this finetune!

Sign up or log in to comment