msmarco-bert-base-dot-v5
This is a sentence-transformers model: It maps sentences & paragraphs to a 768 dimensional dense vector space and was designed for semantic search. It has been trained on 500K (query, answer) pairs from the MS MARCO dataset. For an introduction to semantic search, have a look at: SBERT.net - Semantic Search
Params
110 M
Context
512
Downloads 30d
557 K
Likes
21
Download history
daily snapshots · 10 days607 K557 K
Jul 28Jul 31Aug 3Aug 6
Can you run it?
Estimated VRAM at 8K context unless noted. Pick your hardware to see the verdict per quantization.
| File | Quant | Size | Est. VRAM | Verdict on RTX 4090 · 24 GB |
|---|---|---|---|---|
| model.safetensors | f32 | 0.4 GB | 1.0 GB | ✅ Runs comfortably |
Estimate: file size × 1.1 + KV cache at 8K + 0.5 GB overhead. Not a benchmark — how we calculate this.
Specifications
- Architecture
- BertModel
- Parameters
- 110 M
- Tensor type
- F32
- Context length
- 512
- Vocabulary
- 30,522
- Layers / heads
- 12 / 12
- First seen on the Hub
- 2022-03-02
- Training datasets
- undisclosed
- Added to our catalog
- 2026-07-28
Compare with
Sponsored · GPU cloud
Not enough VRAM?
Spin up a 24 GB L4 instance in 40 seconds. $0.44/hr.