Sentence Similarity
sentence-transformers
Safetensors
Transformers
bert
feature-extraction
biology
protein language model
text-embeddings-inference
Instructions to use monsoon-nlp/protein-matryoshka-embeddings with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- sentence-transformers
How to use monsoon-nlp/protein-matryoshka-embeddings with sentence-transformers:
from sentence_transformers import SentenceTransformer model = SentenceTransformer("monsoon-nlp/protein-matryoshka-embeddings") sentences = [ "That is a happy person", "That is a happy dog", "That is a very happy person", "Today is a sunny day" ] embeddings = model.encode(sentences) similarities = model.similarity(embeddings, embeddings) print(similarities.shape) # [4, 4] - Transformers
How to use monsoon-nlp/protein-matryoshka-embeddings with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoTokenizer, AutoModel tokenizer = AutoTokenizer.from_pretrained("monsoon-nlp/protein-matryoshka-embeddings") model = AutoModel.from_pretrained("monsoon-nlp/protein-matryoshka-embeddings", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -8,9 +8,9 @@ tags:
|
|
| 8 |
- sentence-similarity
|
| 9 |
- transformers
|
| 10 |
- biology
|
|
|
|
| 11 |
license: cc
|
| 12 |
base_model: Rostlab/prot_bert_bfd
|
| 13 |
-
|
| 14 |
---
|
| 15 |
|
| 16 |
# Protein Matryoshka Embeddings
|
|
@@ -87,4 +87,4 @@ This page will be updated when I have examples using it on protein classificatio
|
|
| 87 |
|
| 88 |
I'm interested in whether [embedding quantization](https://huggingface.co/blog/embedding-quantization) could be even more efficient.
|
| 89 |
|
| 90 |
-
If you want to collaborate on future projects / have resources to train longer on more embeddings, please get in touch.
|
|
|
|
| 8 |
- sentence-similarity
|
| 9 |
- transformers
|
| 10 |
- biology
|
| 11 |
+
- protein language model
|
| 12 |
license: cc
|
| 13 |
base_model: Rostlab/prot_bert_bfd
|
|
|
|
| 14 |
---
|
| 15 |
|
| 16 |
# Protein Matryoshka Embeddings
|
|
|
|
| 87 |
|
| 88 |
I'm interested in whether [embedding quantization](https://huggingface.co/blog/embedding-quantization) could be even more efficient.
|
| 89 |
|
| 90 |
+
If you want to collaborate on future projects / have resources to train longer on more embeddings, please get in touch.
|