Spaces:
Running on Zero
Running on Zero
|
Download README.md from ForeverBlue/GRACE-VLM: direct link, hf CLI and curl.
- Browser
- Download file 1.1 kB
-
https://huggingface.co/spaces/ForeverBlue/GRACE-VLM/resolve/e5f230163523f1d5d5da7a0a1d32f21c65a9e62c/README.md
- Command line
-
hf download hf://spaces/ForeverBlue/GRACE-VLM@e5f230163523f1d5d5da7a0a1d32f21c65a9e62c/README.md
-
curl -L -o README.md https://huggingface.co/spaces/ForeverBlue/GRACE-VLM/resolve/e5f230163523f1d5d5da7a0a1d32f21c65a9e62c/README.md
1.1 kB
metadata
title: GRACE-VLM
emoji: 🦢
colorFrom: blue
colorTo: purple
sdk: gradio
python_version: '3.10'
sdk_version: 5.49.1
app_file: app.py
pinned: true
short_description: Try GRACE-VLM on an image, then deploy the INT4 build.
suggested_hardware: zero-a10g
startup_duration_timeout: 1h
preload_from_hub:
- ForeverBlue/Qwen3-VL-2B-GRACE-BF16
models:
- ForeverBlue/Qwen3-VL-2B-GRACE-W4G128-AWQ
- ForeverBlue/Qwen3-VL-2B-GRACE-BF16
tags:
- vision-language-model
- multimodal
- int4
- awq
- knowledge-distillation
- arxiv:2601.22709
GRACE-VLM
Live demo for GRACE-VLM: INT4 Quantization-Aware Distillation for Vision-Language Models, accepted at ICML 2026. Read the paper at arXiv:2601.22709.
Upload an image and ask a question to run the GRACE 2B BF16 checkpoint on free
ZeroGPU hardware. For genuine packed INT4 inference, use
ForeverBlue/Qwen3-VL-2B-GRACE-W4G128-AWQ
with the copy-ready GRACE loader.