NVDA 223.96 ▲2.27%GOOGL 354.30 ▼0.96%MSFT 499.99 ▲0.03%AMD 483.36 ▼1.21%INTC 101.65 ▲1.84%TSMC 420.04 ▲0.44%AMZN 274.48 ▲0.82%META 592.10 ▲0.37%AAPL 313.33 ▲0.29%PLTR 172.01 ▲10.32%
Markets at last close

Models

OpenRouter ranks leading coding models by developer usage

·1 min read

OpenRouter’s coding model rankings, updated July 2026, compare LLMs by real developer usage across code generation, debugging, refactoring, and AI coding assistant workflows. The leaderboard places Xiaomi’s Mimo V2.5 first with 7.54T tokens and 29.1% usage share, followed by MiniMax M3 with 2.79T and 10.8%, Tencent’s Hy3 (free) with 2.52T and 9.7%, Z.ai’s GLM 5.2 with 2.25T and 8.7%, and NVIDIA’s Nemotron 3 Ultra 550B A55B (free) with 1.98T and 7.6%.

The collection emphasizes long-context, agentic coding systems. Several leading models support a 1M-token context window, including MiMo-V2.5, MiniMax-M3, GLM 5.2, DeepSeek V4 Flash, DeepSeek V4 Pro, and Claude Opus 4.8. DeepSeek V4 Flash is positioned for fast inference and high-throughput workloads, while DeepSeek V4 Pro targets more complex tasks such as full-codebase analysis, multi-step automation, and large-scale information synthesis.

Model profiles highlight a mix of multimodal inputs, selectable reasoning levels, and cost-conscious inference. Hy3 offers a 256K context window and configurable reasoning modes, Claude Opus 4.8 is framed for autonomous agents and long-running project orchestration, and Step 3.7 Flash combines image and video understanding with selectable reasoning levels for coding, structured outputs, and long-context productivity tasks.

Originally reported by openrouter.aiRead the source →
Related coverage