Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

CATIE-AQ
/
SPLADE_camembert-base_STS

Feature Extraction
sentence-transformers
Safetensors
French
camembert
sparse-encoder
sparse
splade
Generated from Trainer
dataset_size:12227
loss:SpladeLoss
loss:SparseCosineSimilarityLoss
loss:FlopsLoss
Eval Results (legacy)
text-embeddings-inference
Model card Files Files and versions
xet
Community

Instructions to use CATIE-AQ/SPLADE_camembert-base_STS with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • sentence-transformers

    How to use CATIE-AQ/SPLADE_camembert-base_STS with sentence-transformers:

    from sentence_transformers import SparseEncoder
    
    model = SparseEncoder("CATIE-AQ/SPLADE_camembert-base_STS")
    
    queries = ["Which planet is known as the Red Planet?"]
    documents = [
    	"Venus is often called Earth's twin because of its similar size and proximity.",
    	"Mars, known for its reddish appearance, is often referred to as the Red Planet.",
    	"Jupiter, the largest planet in our solar system, has a prominent red spot.",
    ]
    
    query_embeddings = model.encode_query(queries)
    document_embeddings = model.encode_document(documents)
    
    similarities = model.similarity(query_embeddings, document_embeddings)
    print(similarities)
  • Notebooks
  • Google Colab
  • Kaggle
SPLADE_camembert-base_STS
446 MB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 3 commits
Loïck
Update README.md
b47348d verified over 1 year ago
  • 1_SpladePooling
    Training complete over 1 year ago
  • .gitattributes
    1.52 kB
    initial commit over 1 year ago
  • README.md
    19.8 kB
    Update README.md over 1 year ago
  • added_tokens.json
    28 Bytes
    Training complete over 1 year ago
  • config.json
    664 Bytes
    Training complete over 1 year ago
  • config_sentence_transformers.json
    277 Bytes
    Training complete over 1 year ago
  • model.safetensors
    443 MB
    xet
    Training complete over 1 year ago
  • modules.json
    274 Bytes
    Training complete over 1 year ago
  • sentence_bert_config.json
    57 Bytes
    Training complete over 1 year ago
  • sentencepiece.bpe.model
    811 kB
    xet
    Training complete over 1 year ago
  • special_tokens_map.json
    374 Bytes
    Training complete over 1 year ago
  • tokenizer.json
    2.42 MB
    Training complete over 1 year ago
  • tokenizer_config.json
    1.79 kB
    Training complete over 1 year ago