Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

tiyuvta
/
Qwen3.8-27B-NVFP4-MTP-GGUF

Text Generation
memra
GGUF
nvfp4
speculative-decoding
mtp
conversational
blackwell
qwen3
Eval Results (legacy)
Model card Files Files and versions
xet
Community
3

Instructions to use tiyuvta/Qwen3.8-27B-NVFP4-MTP-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • memra

    How to use tiyuvta/Qwen3.8-27B-NVFP4-MTP-GGUF with memra:

    # memra serves NVIDIA Blackwell workstation and consumer cards (sm_120a), with a
    # compile-gated Hopper lane. Prebuilt binaries need Linux x86_64 and driver 580+,
    # and no CUDA toolkit.
    curl -fsSL https://raw.githubusercontent.com/avifenesh/memra/main/tools/install.sh | sh
    # One chat-templated generation. In a repo with several GGUF files, append
    # :<substring> to choose one, for example hf:tiyuvta/Qwen3.8-27B-NVFP4-MTP-GGUF:Q4_K_M
    MEMRA_CHAT=1 run-gen hf:tiyuvta/Qwen3.8-27B-NVFP4-MTP-GGUF --prompt "Explain KV caches in one sentence."
    # Or an OpenAI-compatible server on 127.0.0.1:8080.
    MEMRA_MODELS="model=hf:tiyuvta/Qwen3.8-27B-NVFP4-MTP-GGUF" memra-server
  • Notebooks
  • Google Colab
  • Kaggle
Qwen3.8-27B-NVFP4-MTP-GGUF
19.4 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 20 commits
Avifenesh's picture
Avifenesh
Point self-references at tiyuvta after org transfer
c303a1b verified 6 days ago
  • .gitattributes
    1.95 kB
    Upload mtp-Qwen3.8-27B-NVFP4-frspec-mixed32768.gguf with huggingface_hub 21 days ago
  • Qwen3.8-27B-NVFP4-Q5K-mtp.gguf
    15.7 GB
    xet
    Upload Qwen3.8-27B-NVFP4-Q5K-mtp.gguf with huggingface_hub 22 days ago
  • README.md
    14.8 kB
    Point self-references at tiyuvta after org transfer 6 days ago
  • mtp-Qwen3.8-27B-NVFP4-frspec-mixed32768.gguf
    1.24 GB
    xet
    Upload mtp-Qwen3.8-27B-NVFP4-frspec-mixed32768.gguf with huggingface_hub 21 days ago
  • mtp-Qwen3.8-27B-NVFP4-frspec-prose32768.gguf
    1.24 GB
    xet
    Upload mtp-Qwen3.8-27B-NVFP4-frspec-prose32768.gguf with huggingface_hub 21 days ago
  • mtp-Qwen3.8-27B-NVFP4-frspec-sxc32768.gguf
    1.24 GB
    xet
    Upload mtp-Qwen3.8-27B-NVFP4-frspec-sxc32768.gguf with huggingface_hub 22 days ago
  • q38-ranks-mixed-32768.txt
    187 kB
    Upload q38-ranks-mixed-32768.txt with huggingface_hub 21 days ago
  • q38-ranks-prose-32768.gguf
    131 kB
    xet
    Upload q38-ranks-prose-32768.gguf with huggingface_hub 21 days ago
  • q38-ranks-prose-32768.txt
    187 kB
    Upload q38-ranks-prose-32768.txt with huggingface_hub 21 days ago
  • q38-ranks-sxc32768.gguf
    131 kB
    xet
    Upload q38-ranks-sxc32768.gguf with huggingface_hub 22 days ago
  • q38-ranks-sxc32768.gguf.txt
    186 kB
    Upload q38-ranks-sxc32768.gguf.txt with huggingface_hub 22 days ago