大模型排行榜

文本嵌入模型排行榜:RAG 选型看这一张

嵌入模型决定了检索质量与成本,请求量高通常意味着稳定、便宜、生态成熟。

数据截止:2026年9月21日 · 本次更新:2026-09-22 21:22(北京时间) · 统计模型数:34

今日文本嵌入榜(Top 50)

最近一个完整自然日文本嵌入模型的请求数排名,变化列为环比。统计窗口内进入榜单的模型共 34 个,合计 5541万 次。

切换窗口: 今日 本周 本月

排名模型厂商请求数环比占比
1 Text Embedding 3 Small OpenAI 2467万 次 +18% 44.5%
2 Qwen3 Embedding 8B 阿里云通义千问 949万 次 +28% 17.1%
3 Gemini Embedding 2 Google 谷歌 464万 次 +40% 8.38%
4 Bge M3 智源研究院 BAAI 383万 次 +54% 6.92%
5 Text Embedding 3 Large OpenAI 352万 次 +61% 6.36%
6 Gemini Embedding 001 Google 谷歌 265万 次 +12% 4.78%
7 Qwen3 Embedding 4B 阿里云通义千问 148万 次 -43% 2.67%
8 Pplx Embed V1 0.6B Perplexity 96.3万 次 -7% 1.74%
9 Gemini Embedding 2 Preview Google 谷歌 77.0万 次 -29% 1.39%
10 All Minilm L12 V2 开源向量模型社区 75.8万 次 +1% 1.37%
11 Voyage 4 Voyage AI 55.2万 次 +196% 0.997%
12 Pplx Embed V1 4B Perplexity 49.3万 次 +8% 0.889%
13 Text Embedding Ada 002 OpenAI 40.8万 次 +21% 0.736%
14 Voyage 4 Large Voyage AI 16.3万 次 +7% 0.294%
15 Voyage Multimodal 3.5 Voyage AI 16.0万 次 +62% 0.289%
16 Multilingual E5 Large 微软 Intfloat 15.9万 次 +6% 0.287%
17 Voyage 4 Lite Voyage AI 14.6万 次 +0% 0.263%
18 Mistral Embed 2312 Mistral AI 13.7万 次 +74% 0.247%
19 Gemini Embedding 2(批处理版) 批处理版 Google 谷歌 10.1万 次 +>999% 0.182%
20 Bge Base En V1.5 智源研究院 BAAI 7.58万 次 +3% 0.137%
21 All Minilm L6 V2 开源向量模型社区 6.45万 次 +103% 0.116%
22 Llama Nemotron Embed VL 1B V2(免费版) 免费版 英伟达 NVIDIA 4.88万 次 +5% 0.088%
23 Nemotron 3 Embed 1B(免费版) 免费版 英伟达 NVIDIA 3.76万 次 +13% 0.068%
24 E5 Large V2 微软 Intfloat 1.41万 次 -48% 0.025%
25 Gte Base thenlper 1.30万 次 +37% 0.023%
26 Bge Large En V1.5 智源研究院 BAAI 1.22万 次 +133% 0.022%
27 Lfm 2.5 Embedding 350M(免费版) 免费版 Liquid AI 1.12万 次 -8% 0.020%
28 Voyage Code 4 Voyage AI 1.11万 次 +192% 0.020%
29 Codestral Embed 2505 Mistral AI 1.08万 次 -7% 0.019%
30 All Mpnet Base V2 开源向量模型社区 1.06万 次 +50% 0.019%
31 Gte Large thenlper 2046 次 -94% 0.004%
32 E5 Base V2 微软 Intfloat 1754 次 -9% 0.003%
33 Multi Qa Mpnet Base Dot V1 开源向量模型社区 20 次 -90% 0.000%
34 Paraphrase Minilm L6 V2 开源向量模型社区 15 次 -97% 0.000%

今日文本嵌入榜厂商份额

同一窗口内按厂商聚合的请求量与份额。

排名厂商上榜模型数请求数份额主力模型
1 OpenAI 3 2860万 次 51.6% Text Embedding 3 Small 44.5%
2 阿里云通义千问 2 1097万 次 19.8% Qwen3 Embedding 8B 17.1%
3 Google 谷歌 4 816万 次 14.7% Gemini Embedding 2 8.38%
4 智源研究院 BAAI 3 392万 次 7.08% Bge M3 6.92%
5 Perplexity 2 146万 次 2.63% Pplx Embed V1 0.6B 1.74%
6 Voyage AI 5 103万 次 1.86% Voyage 4 0.997%
7 开源向量模型社区 5 83.3万 次 1.50% All Minilm L12 V2 1.37%
8 微软 Intfloat 3 17.5万 次 0.316% Multilingual E5 Large 0.287%
9 Mistral AI 2 14.8万 次 0.267% Mistral Embed 2312 0.247%
10 英伟达 NVIDIA 2 8.64万 次 0.156% Llama Nemotron Embed VL 1B V2(免费版) 0.088%
11 thenlper 2 1.51万 次 0.027% Gte Base 0.023%
12 Liquid AI 1 1.12万 次 0.020% Lfm 2.5 Embedding 350M(免费版) 0.020%

近半年请求量走势

按周汇总的请求量走势,取当前请求量最高的 6 个模型,用于判断谁在持续上升、谁在回落。

03908万7816万1.17亿1.56亿tokens / 周3/304/275/256/227/208/179/14Text Embedding 3 SmallQwen3 Embedding 8BBge M3Gemini Embedding 2Text Embedding 3 LargeGemini Embedding 001

这张榜怎么读

常见问题

嵌入模型按什么排名?

按模型处理的嵌入请求数排名,反映被检索/索引任务的调用规模。

向量维度与价格应该怎么看?

维度影响存储成本,价格影响索引预算;两者要结合检索效果一起评估,建议用你的语料做小规模离线对比。

批处理版为什么单独排名?

批处理更适合离线建索引,价格更低但延迟高,因此作为独立变体统计与排名。