|
Download README.md from position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm: direct link, hf CLI and curl.
- Browser
- Download file 374 Bytes
-
https://huggingface.co/position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/resolve/main/README.md
- Command line
-
hf download hf://position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/README.md
-
curl -L -o README.md https://huggingface.co/position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/resolve/main/README.md
374 Bytes
SPEED: Specialized Position Experts for Efficient Speculative Decoding
This repository provides the model checkpoint for SPEED, a speculative decoding method proposed in the anonymous paper:
SPEED: Specialized Position Experts for Efficient Speculative Decoding
📦 Files
pytorch_model.bin— model weightsconfig.json— model configuration