How to use from
vLLM
Install from pip and serve model
# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "aixk/Qwen2.5-Coder-0.5B-Instruct-GGUF"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "aixk/Qwen2.5-Coder-0.5B-Instruct-GGUF",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Use Docker
docker model run hf.co/aixk/Qwen2.5-Coder-0.5B-Instruct-GGUF:
Quick Links
ISAI Logo

ISAI - The Integrated AI Service Platform

ISAI is a comprehensive AI portal offering a variety of practical artificial intelligence services for everyday life.
Discover our diverse family sites and services that enhance convenience and create new digital experiences.

ISAI link ollapp link Addly link blogig link
logig link AI Magician link 99s link Global Stock link
AI Archive link wikiwi link wwwiki link Oduck link
lai link spirit browser link 799 link thedeouk link
wallpaper forum link webbar link Stode link OMAP link
hummorabbit link ollone link ranovel.kr link adsense forum link
Downloads last month
157
GGUF
Model size
0.5B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

1-bit

2-bit

3-bit

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including aixk/Qwen2.5-Coder-0.5B-Instruct-GGUF