Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

CATIE-AQ
/
SPLADE_camemberta2.0_STS

Feature Extraction
sentence-transformers
Safetensors
French
deberta-v2
sparse-encoder
sparse
csr
Generated from Trainer
dataset_size:12227
loss:SpladeLoss
loss:SparseCosineSimilarityLoss
loss:FlopsLoss
Eval Results (legacy)
text-embeddings-inference
Model card Files Files and versions
xet
Community

Instructions to use CATIE-AQ/SPLADE_camemberta2.0_STS with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • sentence-transformers

    How to use CATIE-AQ/SPLADE_camemberta2.0_STS with sentence-transformers:

    from sentence_transformers import SparseEncoder
    
    model = SparseEncoder("CATIE-AQ/SPLADE_camemberta2.0_STS")
    
    queries = ["Which planet is known as the Red Planet?"]
    documents = [
    	"Venus is often called Earth's twin because of its similar size and proximity.",
    	"Mars, known for its reddish appearance, is often referred to as the Red Planet.",
    	"Jupiter, the largest planet in our solar system, has a prominent red spot.",
    ]
    
    query_embeddings = model.encode_query(queries)
    document_embeddings = model.encode_document(documents)
    
    similarities = model.similarity(query_embeddings, document_embeddings)
    print(similarities)
  • Notebooks
  • Google Colab
  • Kaggle
SPLADE_camemberta2.0_STS
453 MB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 2 commits
Loïck
Training complete
cded0c2 verified over 1 year ago
  • 1_Pooling
    Training complete over 1 year ago
  • 2_SparseAutoEncoder
    Training complete over 1 year ago
  • .gitattributes
    1.52 kB
    initial commit over 1 year ago
  • README.md
    20.1 kB
    Training complete over 1 year ago
  • config.json
    981 Bytes
    Training complete over 1 year ago
  • config_sentence_transformers.json
    277 Bytes
    Training complete over 1 year ago
  • model.safetensors
    442 MB
    xet
    Training complete over 1 year ago
  • modules.json
    380 Bytes
    Training complete over 1 year ago
  • sentence_bert_config.json
    58 Bytes
    Training complete over 1 year ago
  • special_tokens_map.json
    971 Bytes
    Training complete over 1 year ago
  • tokenizer.json
    756 kB
    Training complete over 1 year ago
  • tokenizer_config.json
    1.28 kB
    Training complete over 1 year ago
  • vocab.txt
    235 kB
    Training complete over 1 year ago