Weekly AI Ranking
The 12 most notable AI products/models this week, ranked by real-world impact.
This week's highlights
- New entries this week: DeepSeek V4, Kimi K3.
- Climbing the ranks: Claude, Cursor, NotebookLM.
- No entries dropped this week.
- 7 products held their position.
Ranking methodology
The npmrundeploy.com team hand-curates this table every week based on social/dev-community discussion volume, product launch news and hands-on trials — not an automated algorithm or benchmark. Ranks reflect the editorial team's judgment at publish time, and may differ from objective benchmark leaderboards like the Intelligence Index or OpenRouter below.
🧠 Model intelligence index
Intelligence Index — Artificial Analysis
An independent ranking of AI model "intelligence," aggregated from a battery of benchmark tests run by Artificial Analysis.
Data source: Artificial Analysis — View source →
📊 Ranked by real-world usage
OpenRouter Rankings — this week's usage volume
Rankings based on real usage data (not benchmarks) from millions of requests flowing through OpenRouter each week, broken down across 7 task categories.
Week of 2026-07-25
🏆 Top models by overall usage
- 1 DeepSeek V4 Flash deepseek 7.22T tokens +13%
- 2 MiMo-V2.5 xiaomi 6.3T tokens +40%
- 3 Hy3 tencent 4.82T tokens +22%
- 4 DeepSeek V4 Pro deepseek 3.28T tokens +3%
- 5 GLM 5.2 z-ai 2.89T tokens +12%
- 6 Nemotron 3 Ultra (free) nvidia 2.43T tokens +4%
- 7 MiniMax M3 minimax 1.96T tokens +4%
- 8 GPT-5.6 Luna openai 1.95T tokens +465%
- 9 Step 3.7 Flash stepfun 1.66T tokens +14%
- 10 Kimi K3 moonshotai 1.42T tokens +16%
⚡ Quick start via OpenRouter
OpenRouter lets you call almost every model on the market through a single API endpoint (OpenAI-compatible), without registering separately with each provider.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "<model-id>", "messages": [{"role":"user","content":"Xin chào"}]}' Text (LLM)
- 1 MiMo-V2.5 9.95T tokens +9%
- 2 DeepSeek V4 Flash 5.81T tokens +9%
- 3 Hy3 (free) 4.93T tokens +53%
- 4 GLM 5.2 3.58T tokens +1%
- 5 DeepSeek V4 Pro 3.01T tokens +16%
Quick take: MiMo-V2.5 (Xiaomi) leads this week by token volume — a fast-rising open model; Tencent's Hy3 (free) jumped 53%, showing strong demand for free-tier trials.
Image
- 1 Nano Banana (Gemini 2.5 Flash Image) 1.3M requests +5%
- 2 Nano Banana 2 (Gemini 3.1 Flash Image) 791K requests +16%
- 3 Nano Banana 2 (Gemini 3.1 Flash Image Preview) 727K requests +11%
- 4 Nano Banana Pro (Gemini 3 Pro Image Preview) 520K requests +0%
- 5 GPT Image 2 515K requests +13%
Quick take: Google's "Nano Banana" family (Gemini 2.5/3.1 Flash Image) takes 4 of the top 5 image-gen spots — a clear distribution edge via Gemini, ahead of OpenAI's GPT Image 2 in 5th.
Embeddings
- 1 Text Embedding 3 Small 87.5M requests +10%
- 2 Qwen3 Embedding 8B 68.7M requests +3%
- 3 bge-m3 16.2M requests +55%
- 4 Text Embedding 3 Large 14.1M requests +21%
- 5 Embed V1 0.6B 12.3M requests +20%
Quick take: OpenAI's Text Embedding 3 Small remains #1 by a wide margin (87.5M requests/week); Qwen3 Embedding 8B holds a close 2nd, while Gemini Embedding variants are the fastest growers in the 6-8 range.
Rerank
- 1 Llama Nemotron Rerank VL 1B V2 (free) 1.49M requests +26%
- 2 Rerank v3.5 689K requests +6%
- 3 Rerank 4 Fast 582K requests +141%
- 4 Rerank 4 Pro 310K requests +10%
Quick take: NVIDIA's free Llama Nemotron Rerank leads by volume, while Cohere completely dominates the rest of the paid top 4 — no other provider breaks in.
Video
- 1 Veo 3.1 Fast 37K requests +17%
- 2 Seedance 2.0 25K requests +13%
- 3 Veo 3.1 Lite 15K requests +35%
- 4 Veo 3.1 13K requests +30%
- 5 Seedance 2.0 Fast 11K requests +28%
Quick take: Google's Veo 3.1 Fast leads video generation, with ByteDance's Seedance 2.0 close behind in 2nd — these two dominate share, and the faster "Fast/Lite" variants outpace the full versions.
Speech (TTS)
- 1 Kokoro 82M 604K requests +78%
- 2 Gemini 3.1 Flash TTS Preview 415K requests +9%
- 3 Grok Voice TTS 1.0 326K requests +95%
- 4 MAI-Voice-2 31K requests +35%
- 5 Aura-2 8K requests +175%
Quick take: Kokoro 82M — a compact open-source TTS model by hexgrad — leads usage; x-AI's Grok Voice TTS is growing fast (+95%), and Deepgram's Aura-2 surged +175% despite still ranking lower.
Transcription
- 1 Whisper Large V3 3.6M requests +238%
- 2 Whisper Large V3 Turbo 1.57M requests +57%
- 3 GPT-4o Mini Transcribe 675K requests +33%
- 4 Qwen3 ASR Flash 247K requests +7%
- 5 Whisper 1 174K requests +54%
Quick take: OpenAI's Whisper family sweeps the top 2 spots; Whisper Large V3 usage spiked +238% this week, showing surging API transcription demand even as challengers like Qwen3 ASR Flash keep pace.
Data source: OpenRouter — View source →
🎯 Ranked by aggregated benchmarks
Onyx LLM Leaderboard — weekly summary
Onyx aggregates many benchmarks (reasoning, coding, math, agentic, SWE-bench, chat) and sorts models into 5 tiers, S/A/B/C/D — S being the clear leaders.
Data as of
Data source: Onyx LLM Leaderboard — View source → · Data as of