Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
AutomatosX
/
AX-Qwen3-Embedding-8B-CUDA-AXQ-NVFP4-W4A4
Like
0
Sentence Similarity
Safetensors
vllm
qwen3
text-embeddings
retrieval
cuda
nvfp4
w4a4
compressed-tensors
axquant
development-preview
8-bit precision
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
AX-Qwen3-Embedding-8B-CUDA-AXQ-NVFP4-W4A4
/
evaluation
3.74 MB
Ctrl+K
Ctrl+K
1 contributor
History:
1 commit
AutomatosX
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
f93bb79
verified
4 days ago
retrieval-corpus.json
Safe
806 Bytes
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
4 days ago
rtx5090-bf16.json
Safe
935 kB
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
4 days ago
rtx5090-nvfp4.json
Safe
935 kB
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
4 days ago
thor-bf16.json
Safe
935 kB
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
4 days ago
thor-nvfp4.json
Safe
935 kB
Add native AXQuant NVFP4 W4A4 development checkpoint with two-GPU evidence
4 days ago