cross-encoder/ms-marco-MiniLM-L-6-v2
Primitive: /score · Score ·
BERT
This model was trained on the MS Marco Passage Ranking task.
View on Hugging Face → Fine-tuned from cross-encoder/ms-marco-MiniLM-L12-v2
Overview
Hardware: — drives latency, throughput & cost
| Size | 23M params |
|---|---|
| Tasks | /score |
| License | apache-2.0 |
| Languages | en |
| Latency | 46 ms |
| Throughput | 51.1K tok/s |
| Cost | $0.0043 /1M tok |
Cost is approximate — computed from list GPU prices; your actual price depends on the provider you deploy SIE with.
Scoring
| Inputs | text |
|---|---|
| Max sequence length | 512 |
Benchmarks
AskUbuntuDupQuestions
Duplicate question detection from AskUbuntu
CMedQAv1Reranking
Chinese medical question answering reranking (v1)
CMedQAv2Reranking
Chinese medical question answering reranking (v2)
CQADupstackPhysicsRetrieval?candidates_model=Alibaba-NLP
CosQA?candidates_model=Alibaba-NLP
FiQA2018?candidates_model=Alibaba-NLP
LegalBenchConsumerContractsQA?candidates_model=Alibaba-NLP
MMarcoReranking
Multilingual MARCO passage reranking (Chinese)
NFCorpus?candidates_model=Alibaba-NLP
NanoFiQA2018Retrieval
Smaller subset of the FiQA financial QA dataset