大模型排行榜

文本嵌入模型排行榜:RAG 选型看这一张

嵌入模型决定了检索质量与成本,请求量高通常意味着稳定、便宜、生态成熟。

数据截止:2026年9月21日 · 本次更新:2026-09-22 21:22(北京时间) · 统计模型数:34

本月文本嵌入榜(Top 50)

最近 30 天文本嵌入模型的请求数排名,变化列为环比。统计窗口内进入榜单的模型共 34 个,合计 13.0亿 次。

切换窗口: 今日 本周 本月

排名模型厂商请求数环比占比
1 Text Embedding 3 Small OpenAI 5.41亿 次 +23% 41.6%
2 Qwen3 Embedding 8B 阿里云通义千问 2.50亿 次 -12% 19.2%
3 Gemini Embedding 2 Google 谷歌 9635万 次 +39% 7.40%
4 Text Embedding 3 Large OpenAI 8952万 次 +33% 6.87%
5 Bge M3 智源研究院 BAAI 8628万 次 +14% 6.62%
6 Gemini Embedding 001 Google 谷歌 6363万 次 -7% 4.88%
7 Pplx Embed V1 0.6B Perplexity 3531万 次 -51% 2.71%
8 Qwen3 Embedding 4B 阿里云通义千问 2454万 次 +14% 1.88%
9 All Minilm L12 V2 开源向量模型社区 2192万 次 +14% 1.68%
10 Gemini Embedding 2 Preview Google 谷歌 1915万 次 +37% 1.47%
11 Text Embedding Ada 002 OpenAI 1427万 次 +10% 1.10%
12 Pplx Embed V1 4B Perplexity 1148万 次 -36% 0.881%
13 Voyage 4 Lite Voyage AI 856万 次 +>999% 0.657%
14 Voyage 4 Voyage AI 825万 次 +>999% 0.633%
15 Voyage 4 Large Voyage AI 533万 次 +295% 0.409%
16 Mistral Embed 2312 Mistral AI 514万 次 -20% 0.394%
17 Multilingual E5 Large 微软 Intfloat 443万 次 -38% 0.340%
18 All Minilm L6 V2 开源向量模型社区 375万 次 -3% 0.288%
19 Bge Base En V1.5 智源研究院 BAAI 278万 次 -57% 0.214%
20 Llama Nemotron Embed VL 1B V2(免费版) 免费版 英伟达 NVIDIA 253万 次 -82% 0.194%
21 Nemotron 3 Embed 1B(免费版) 免费版 英伟达 NVIDIA 231万 次 -74% 0.177%
22 Voyage Multimodal 3.5 Voyage AI 194万 次 +>999% 0.149%
23 E5 Large V2 微软 Intfloat 99.4万 次 +31% 0.076%
24 Lfm 2.5 Embedding 350M(免费版) 免费版 Liquid AI 71.1万 次 +430% 0.055%
25 Codestral Embed 2505 Mistral AI 60.9万 次 -61% 0.047%
26 Bge Large En V1.5 智源研究院 BAAI 43.1万 次 +71% 0.033%
27 Gte Base thenlper 32.1万 次 -65% 0.025%
28 Voyage Code 4 Voyage AI 28.6万 次 +640% 0.022%
29 All Mpnet Base V2 开源向量模型社区 25.3万 次 +30% 0.019%
30 Gemini Embedding 2(批处理版) 批处理版 Google 谷歌 10.8万 次 新上榜 0.008%
31 Gte Large thenlper 6.00万 次 +86% 0.005%
32 E5 Base V2 微软 Intfloat 3.85万 次 -9% 0.003%
33 Paraphrase Minilm L6 V2 开源向量模型社区 3.84万 次 -47% 0.003%
34 Multi Qa Mpnet Base Dot V1 开源向量模型社区 2126 次 -84% 0.000%

本月文本嵌入榜厂商份额

同一窗口内按厂商聚合的请求量与份额。

排名厂商上榜模型数请求数份额主力模型
1 OpenAI 3 6.45亿 次 49.5% Text Embedding 3 Small 41.6%
2 阿里云通义千问 2 2.75亿 次 21.1% Qwen3 Embedding 8B 19.2%
3 Google 谷歌 4 1.79亿 次 13.8% Gemini Embedding 2 7.40%
4 智源研究院 BAAI 3 8949万 次 6.87% Bge M3 6.62%
5 Perplexity 2 4679万 次 3.59% Pplx Embed V1 0.6B 2.71%
6 开源向量模型社区 5 2596万 次 1.99% All Minilm L12 V2 1.68%
7 Voyage AI 5 2437万 次 1.87% Voyage 4 Lite 0.657%
8 Mistral AI 2 575万 次 0.441% Mistral Embed 2312 0.394%
9 微软 Intfloat 3 546万 次 0.419% Multilingual E5 Large 0.340%
10 英伟达 NVIDIA 2 484万 次 0.371% Llama Nemotron Embed VL 1B V2(免费版) 0.194%
11 Liquid AI 1 71.1万 次 0.055% Lfm 2.5 Embedding 350M(免费版) 0.055%
12 thenlper 2 38.1万 次 0.029% Gte Base 0.025%

近半年请求量走势

按周汇总的请求量走势,取当前请求量最高的 6 个模型,用于判断谁在持续上升、谁在回落。

03908万7816万1.17亿1.56亿tokens / 周3/304/275/256/227/208/179/14Text Embedding 3 SmallQwen3 Embedding 8BBge M3Gemini Embedding 2Text Embedding 3 LargeGemini Embedding 001

这张榜怎么读

常见问题

嵌入模型按什么排名?

按模型处理的嵌入请求数排名,反映被检索/索引任务的调用规模。

向量维度与价格应该怎么看?

维度影响存储成本,价格影响索引预算;两者要结合检索效果一起评估,建议用你的语料做小规模离线对比。

批处理版为什么单独排名?

批处理更适合离线建索引,价格更低但延迟高,因此作为独立变体统计与排名。