How to use from
vLLM
Install from pip and serve model
# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "britllm/britllm-3b-v0.1"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "britllm/britllm-3b-v0.1",
		"prompt": "Once upon a time,",
		"max_tokens": 512,
		"temperature": 0.5
	}'
Use Docker
docker model run hf.co/britllm/britllm-3b-v0.1
Quick Links

This is a raw, pretrained model, which should be further finetuned for most use cases.

Visit our webpage for detailed information: https://llm.org.uk .

Contact
Email: nlp-britllm@cs.ucl.ac.uk

Acknowledgements
We would like to acknowledge the support of DiRAC (Distributed Research using Advanced Computing), Microsoft Research's Accelerate Foundation Models Research Grant, the UCL Centre for Artificial Intelligence, and the Generative Models AI Hub.

© BritLLM

Downloads last month
85
Safetensors
Model size
3B params
Tensor type
F16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for britllm/britllm-3b-v0.1

Finetunes
1 model
Quantizations
1 model