cross-encoder/ms-marco-MiniLM-L-12-v2
Primitive: /score · Score ·
BERT
This model was trained on the MS Marco Passage Ranking task.
View on Hugging Face → Fine-tuned from microsoft/MiniLM-L12-H384-uncased
Overview
Hardware: — drives latency, throughput & cost
| Size | 33M params |
|---|---|
| Tasks | /score |
| License | apache-2.0 |
| Languages | en |
| Latency | 40 ms |
| Throughput | 26.4K tok/s |
| Cost | $0.0084 /1M tok |
Cost is approximate — computed from list GPU prices; your actual price depends on the provider you deploy SIE with.
Scoring
| Inputs | text |
|---|---|
| Max sequence length | 512 |
Benchmarks
AskUbuntuDupQuestions
Duplicate question detection from AskUbuntu
CMedQAv1Reranking
Chinese medical question answering reranking (v1)
CMedQAv2Reranking
Chinese medical question answering reranking (v2)
MMarcoReranking
Multilingual MARCO passage reranking (Chinese)