|
Download README.md from position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm: direct link, hf CLI and curl.
- Browser
- Download file 374 Bytes
-
https://huggingface.co/position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/resolve/main/README.md
- Command line
-
hf download hf://position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/README.md
-
curl -L -o README.md https://huggingface.co/position-specialist-speculative-decoding/Speed-E3-Llama3.1-8B-Instruct-vllm/resolve/main/README.md
374 Bytes
| # SPEED: Specialized Position Experts for Efficient Speculative Decoding | |
| This repository provides the model checkpoint for **SPEED**, a speculative decoding method proposed in the anonymous paper: | |
| > **SPEED: Specialized Position Experts for Efficient Speculative Decoding** | |
| ## 📦 Files | |
| - `pytorch_model.bin` — model weights | |
| - `config.json` — model configuration | |