MODEL · BAAI · 568M (XLM-ROBERTA-LARGE BASE)
BGE-M3
BAAI's multi-functionality + multilingual (170+ languages) + multi-granularity embedding. The default "just use it" RAG embedding since early 2024. As of 2026 it is no longer the top-quality pick — `Qwen3-Embedding` (0.6B / 4B / 8B, Apache 2.0) now leads MTEB overall — but BGE-M3 remains the sharpest pick for cheap, broad multilingual breadth at 568M.
License: MIT · Context: 8192 tokens · Released: February 2024
The decision in five lines
- The call
- Consider — runnable locally, family reference
- Best for
- Local evaluation and family reference
- Runs on
- 23 hardware picks fit (cheapest: Intel Arc B580 12 GB · $310)
- Watch out
- Max-quality general/English retrieval — Qwen3-Embedding-8B (Apache 2.0) ranks #1 on MTEB and is the better default when you have the VRAM.
- Evidence
- Estimated
- 568M (XLM-RoBERTa-large base)
- PARAMETERS
- EMBEDDING
- TYPE
- 8192
- CONTEXT
- ~1–2 GB
- VRAM AT Q4
Where we recommend this
This model isn’t currently in an active planner slot. See the runner notes below if you’re running it anyway.
The call
BAAI's multi-functionality + multilingual (170+ languages) + multi-granularity embedding. The default "just use it" RAG embedding since early 2024. As of 2026 it is no longer the top-quality pick — `Qwen3-Embedding` (0.6B / 4B / 8B, Apache 2.0) now leads MTEB overall — but BGE-M3 remains the sharpest pick for cheap, broad multilingual breadth at 568M.
When not to use: Max-quality general/English retrieval — Qwen3-Embedding-8B (Apache 2.0) ranks #1 on MTEB and is the better default when you have the VRAM. Use BGE-M3 when you want the smallest model that still covers 170+ languages, or a CPU-friendly 568M retriever.
Runner notes
Ollama tag `bge-m3`. Also natively in FlagEmbedding, sentence-transformers, llama.cpp. 568M params run fine on CPU for small corpora. For top quality step up to `Qwen/Qwen3-Embedding-8B` (or the 0.6B/4B for lighter rigs).
Hardware that fits
Every hardware pick whose memory fits this model at the quant we recommend. Sorted cheapest-first — the top row is your best-value fit. Click through for the full buyer’s guide.
- Intel Arc B580 12 GBPerfect · 5.9× 12 GB · $310–$429 new / $190–$260 refurb
- Minisforum UM890 ProPerfect · 11.8× 32 GB DDR5 (shared) · $463 barebone (no RAM / SSD); $991 with 32 GB / 1 TB — Minisforum US store, Sep 15 2026
- NVIDIA RTX 3060 12 GBPerfect · 5.9× 12 GB · $480–$660
- RTX 5060 Ti 16 GBPerfect · 7.9× 16 GB · $700 (refurb / open box) – $970 (new)
- AMD Radeon RX 9070 XTPerfect · 7.9× 16 GB · $750–$870
- Mac Mini M4 16 GBPerfect · 5.3× 16 GB unified · $899 (M6 successor, 16 GB / 256 GB, pre-order) / M4 residuals $499–$799
- NVIDIA RTX 3090 (used, single)Perfect · 11.8× 24 GB · $950–$1,200
- NVIDIA RTX 5070 TiPerfect · 7.9× 16 GB · $1,150–$1,300 new; open-box / refurb from ~$1,020
- AMD Radeon RX 7900 XTXPerfect · 11.8× 24 GB · $1,300–$1,500 new (no refurb in stock Sep 15 2026)
- NVIDIA RTX 5080Perfect · 7.9× 16 GB · $1,430–$1,700
- MacBook Air M5 24 GBPerfect · 7.9× 24 GB unified · $1,499–$1,899
- Mac Mini M4 Pro 24 GBPerfect · 7.9× 24 GB unified · $1,699 (M5 Pro successor, 24 GB / 512 GB, pre-order; M4 Pro discontinued Aug 25 2026)
- Dual RTX 3090 (used)Perfect · 23.6× 48 GB · $1,800–$2,500 all-in
- NVIDIA RTX 4090Perfect · 11.8× 24 GB · $2,200–$2,800
- M5 Pro MacBook Pro 48 GBPerfect · 15.8× 48 GB unified · $2,999–$3,599
- Framework Desktop (Ryzen AI Max+ 395)Perfect · 42.1× 128 GB unified · $3,449 (128 GB config)
- NVIDIA RTX A6000 (48 GB, used)Perfect · 23.6× 48 GB ECC · $3,500–$4,500
- Mac Studio M4 Max 64 GBPerfect · 21.0× 64 GB unified · $3,799 (M5 Max successor, 64 GB / 1 TB, pre-order; M4 Max discontinued Aug 25 2026)
- NVIDIA DGX SparkPerfect · 42.1× 128 GB unified · $4,699
- M5 Max MacBook Pro 64 GBPerfect · 21.0× 64 GB unified · ~$5,199 (est.; June 25 2026 increase)
- Mac Studio M3 Ultra 96 GBPerfect · 31.6× 96 GB unified · $5,499 (M5 Ultra successor, 96 GB / 1 TB, pre-order; M3 Ultra discontinued Aug 25 2026)
- NVIDIA RTX 5090Perfect · 15.7× 32 GB · $6,450–$7,000 (new, in stock)
- Dual RTX 5090Perfect · 31.4× 64 GB (2×32) · $13,300–$14,000 all-in (two cards alone are ~$12,900 at the Sep 15 2026 Newegg floor)
Next step
Find-by-model — see what hardware runs this→