Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

multi-vector-encoder-testing
/
bert-tiny-multi-vector

Feature Extraction
sentence-transformers
Safetensors
English
bert
multi-vector
colbert
late-interaction
Generated from Trainer
dataset_size:501907
loss:MultiVectorMultipleNegativesRankingLoss
Eval Results (legacy)
text-embeddings-inference
Model card Files Files and versions
xet
Community

Instructions to use multi-vector-encoder-testing/bert-tiny-multi-vector with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • sentence-transformers

    How to use multi-vector-encoder-testing/bert-tiny-multi-vector with sentence-transformers:

    from sentence_transformers import MultiVectorEncoder
    
    model = MultiVectorEncoder("multi-vector-encoder-testing/bert-tiny-multi-vector")
    
    queries = ["Which planet is known as the Red Planet?"]
    documents = [
    	"Venus is often called Earth's twin because of its similar size and proximity.",
    	"Mars, known for its reddish appearance, is often referred to as the Red Planet.",
    	"Jupiter, the largest planet in our solar system, has a prominent red spot.",
    ]
    
    query_embeddings = model.encode_query(queries)
    document_embeddings = model.encode_document(documents)
    
    similarities = model.similarity(query_embeddings, document_embeddings)
    print(similarities)
  • Notebooks
  • Google Colab
  • Kaggle
bert-tiny-multi-vector
18.6 MB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 3 commits
tomaarsen's picture
tomaarsen HF Staff
Clean up model naming and remove training comparison
0bcd821 verified 20 days ago
  • 1_Dense
    Add new MultiVectorEncoder model 20 days ago
  • 2_MultiVectorMask
    Add new MultiVectorEncoder model 20 days ago
  • 3_Normalize
    Add new MultiVectorEncoder model 20 days ago
  • .gitattributes
    1.52 kB
    initial commit 20 days ago
  • README.md
    156 kB
    Clean up model naming and remove training comparison 20 days ago
  • config.json
    696 Bytes
    Add new MultiVectorEncoder model 20 days ago
  • config_sentence_transformers.json
    301 Bytes
    Add new MultiVectorEncoder model 20 days ago
  • model.safetensors
    17.5 MB
    xet
    Add new MultiVectorEncoder model 20 days ago
  • modules.json
    588 Bytes
    Add new MultiVectorEncoder model 20 days ago
  • results.json
    89.2 kB
    Add new MultiVectorEncoder model 20 days ago
  • sentence_bert_config.json
    438 Bytes
    Add new MultiVectorEncoder model 20 days ago
  • tokenizer.json
    712 kB
    Add new MultiVectorEncoder model 20 days ago
  • tokenizer_config.json
    564 Bytes
    Add new MultiVectorEncoder model 20 days ago
  • train.py
    9.39 kB
    Clean up model naming and remove training comparison 20 days ago