Text embeddings.
Embedding models convert text into fixed-length vectors that capture semantic meaning. The bedrock of every RAG pipeline, vector database, and semantic-search system. Most embedding models are tiny enough to run on consumer GPUs or even CPU.
Top GPUs for this workload.
Ranked by suitability — higher fitness scores mean the card handles this workload more comfortably.
No GPU recommendations linked to this workload yet.
Top models for this workload.
Nomic Embed Text
Open embedding model — 69M Ollama pulls, the local default.
mxbai-embed-large
335M embedding model — top MTEB scores for its size.
BGE-M3
Multilingual + multifunctional embedding (100+ languages).