Transformers
Safetensors
llama
speculative-decoding
eagle3
specforge
sglang
qwen3
draft-model
sharegpt
sliding-window-512
text-generation-inference
Instructions to use huluhuluu/qwen3-4b-instruct-2507-eagle3-sharegpt-sw512-epoch2-step60000 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use huluhuluu/qwen3-4b-instruct-2507-eagle3-sharegpt-sw512-epoch2-step60000 with Transformers:
# Load model directly from transformers import AutoTokenizer, LlamaForCausalLMEagle3 tokenizer = AutoTokenizer.from_pretrained("huluhuluu/qwen3-4b-instruct-2507-eagle3-sharegpt-sw512-epoch2-step60000") model = LlamaForCausalLMEagle3.from_pretrained("huluhuluu/qwen3-4b-instruct-2507-eagle3-sharegpt-sw512-epoch2-step60000", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Xet hash:
- 3220ecc4502cad03b1994e977a87df0bfa683f922b4981c3ea38b918568c2697
- Size of remote file:
- 4.31 kB
- SHA256:
- 8ff1abcac122963865bc9e96c667ee2dad7b48c6600c95b3e1d541a570c16ce0
·
Xet efficiently stores Large Files inside Git, intelligently splitting files into unique chunks and accelerating uploads and downloads. More info.