大模型排行榜

文本嵌入模型排行榜:RAG 选型看这一张

嵌入模型决定了检索质量与成本,请求量高通常意味着稳定、便宜、生态成熟。

数据截止:2026年9月21日 · 本次更新:2026-09-22 21:22(北京时间) · 统计模型数:34

本周文本嵌入榜(Top 50)

最近 7 天文本嵌入模型的请求数排名,变化列为环比。统计窗口内进入榜单的模型共 34 个,合计 3.49亿 次。

切换窗口: 今日 本周 本月

排名模型厂商请求数环比占比
1 Text Embedding 3 Small OpenAI 1.50亿 次 +22% 43.1%
2 Qwen3 Embedding 8B 阿里云通义千问 5985万 次 +5% 17.2%
3 Gemini Embedding 2 Google 谷歌 2665万 次 +26% 7.64%
4 Bge M3 智源研究院 BAAI 2258万 次 +11% 6.48%
5 Text Embedding 3 Large OpenAI 1963万 次 -3% 5.63%
6 Gemini Embedding 001 Google 谷歌 1931万 次 +5% 5.54%
7 Qwen3 Embedding 4B 阿里云通义千问 795万 次 +28% 2.28%
8 Gemini Embedding 2 Preview Google 谷歌 762万 次 +87% 2.19%
9 Pplx Embed V1 0.6B Perplexity 744万 次 -10% 2.13%
10 Voyage 4 Lite Voyage AI 619万 次 +398% 1.77%
11 All Minilm L12 V2 开源向量模型社区 513万 次 +117% 1.47%
12 Pplx Embed V1 4B Perplexity 403万 次 +22% 1.15%
13 Voyage 4 Voyage AI 265万 次 +31% 0.759%
14 Text Embedding Ada 002 OpenAI 234万 次 -23% 0.672%
15 Multilingual E5 Large 微软 Intfloat 123万 次 +14% 0.352%
16 Mistral Embed 2312 Mistral AI 119万 次 +22% 0.342%
17 Voyage 4 Large Voyage AI 107万 次 +11% 0.307%
18 Voyage Multimodal 3.5 Voyage AI 71.0万 次 -10% 0.204%
19 All Minilm L6 V2 开源向量模型社区 60.2万 次 -40% 0.173%
20 Bge Base En V1.5 智源研究院 BAAI 54.7万 次 -10% 0.157%
21 Llama Nemotron Embed VL 1B V2(免费版) 免费版 英伟达 NVIDIA 35.1万 次 -7% 0.101%
22 Nemotron 3 Embed 1B(免费版) 免费版 英伟达 NVIDIA 24.9万 次 -4% 0.071%
23 Bge Large En V1.5 智源研究院 BAAI 19.4万 次 +262% 0.056%
24 E5 Large V2 微软 Intfloat 17.8万 次 -27% 0.051%
25 Codestral Embed 2505 Mistral AI 12.7万 次 +38% 0.036%
26 Gemini Embedding 2(批处理版) 批处理版 Google 谷歌 10.8万 次 新上榜 0.031%
27 Lfm 2.5 Embedding 350M(免费版) 免费版 Liquid AI 8.32万 次 -6% 0.024%
28 Gte Base thenlper 6.60万 次 +23% 0.019%
29 All Mpnet Base V2 开源向量模型社区 6.24万 次 +10% 0.018%
30 Voyage Code 4 Voyage AI 5.66万 次 -24% 0.016%
31 Gte Large thenlper 4.32万 次 +>999% 0.012%
32 E5 Base V2 微软 Intfloat 1.11万 次 +114% 0.003%
33 Paraphrase Minilm L6 V2 开源向量模型社区 1550 次 -90% 0.000%
34 Multi Qa Mpnet Base Dot V1 开源向量模型社区 531 次 -27% 0.000%

本周文本嵌入榜厂商份额

同一窗口内按厂商聚合的请求量与份额。

排名厂商上榜模型数请求数份额主力模型
1 OpenAI 3 1.72亿 次 49.4% Text Embedding 3 Small 43.1%
2 阿里云通义千问 2 6780万 次 19.4% Qwen3 Embedding 8B 17.2%
3 Google 谷歌 4 5368万 次 15.4% Gemini Embedding 2 7.64%
4 智源研究院 BAAI 3 2332万 次 6.69% Bge M3 6.48%
5 Perplexity 2 1147万 次 3.29% Pplx Embed V1 0.6B 2.13%
6 Voyage AI 5 1067万 次 3.06% Voyage 4 Lite 1.77%
7 开源向量模型社区 5 580万 次 1.66% All Minilm L12 V2 1.47%
8 微软 Intfloat 3 142万 次 0.406% Multilingual E5 Large 0.352%
9 Mistral AI 2 132万 次 0.378% Mistral Embed 2312 0.342%
10 英伟达 NVIDIA 2 60.0万 次 0.172% Llama Nemotron Embed VL 1B V2(免费版) 0.101%
11 thenlper 2 10.9万 次 0.031% Gte Base 0.019%
12 Liquid AI 1 8.32万 次 0.024% Lfm 2.5 Embedding 350M(免费版) 0.024%

近半年请求量走势

按周汇总的请求量走势,取当前请求量最高的 6 个模型,用于判断谁在持续上升、谁在回落。

03908万7816万1.17亿1.56亿tokens / 周3/304/275/256/227/208/179/14Text Embedding 3 SmallQwen3 Embedding 8BBge M3Gemini Embedding 2Text Embedding 3 LargeGemini Embedding 001

这张榜怎么读

常见问题

嵌入模型按什么排名?

按模型处理的嵌入请求数排名,反映被检索/索引任务的调用规模。

向量维度与价格应该怎么看?

维度影响存储成本,价格影响索引预算;两者要结合检索效果一起评估,建议用你的语料做小规模离线对比。

批处理版为什么单独排名?

批处理更适合离线建索引,价格更低但延迟高,因此作为独立变体统计与排名。