NVDA 228.87 ▲0.66%GOOGL 351.16 ▼1.07%MSFT 498.00 ▼0.72%AMD 623.77 ▲1.34%INTC 123.86 ▲1.71%TSMC 452.00 ▲1.54%AMZN 254.98 ▼1.34%META 736.60 ▼0.63%AAPL 339.75 ▲0.23%PLTR 184.99 ▲1.04%
Markets at last close

OpenAI · Models

LLM API prices show wide gap across providers

·1 min read

Pricing verified June 28, 2026 shows a wide gap among hosted LLM APIs. DeepSeek V4 Flash is listed as the cheapest option at $0.14/$0.28 per 1M tokens, while GPT-5.5 is priced at $5/$30 and Claude Opus 4.8 at $5/$25. MiniMax M3 at $0.60/$2.40 is framed as the strongest value pick, with a score above 80% on SWE-bench Verified and a 1M context.

Quality rankings are concentrated near the top for coding tasks. GPT-5.5 scores 88.7% on SWE-bench Verified and Claude Opus 4.8 scores 88.6%, while Anthropic’s higher-scoring Fable 5 at 95.0% was suspended in June 2026. Several lower-cost models, including DeepSeek V4 Pro Max, Gemini 3.1 Pro, MiniMax M3, Qwen3.7 Max, and Kimi K2.6, fall between 80.2% and 80.6%.

Operational differences are as important as headline rates. Anthropic uses spend-based tiers and cache-aware limits, OpenAI unlocks tiers by payment history, and DeepSeek publishes concurrency caps of 2,500 on V4 Flash and 500 on V4 Pro. Most providers support OpenAI-format requests, and the guidance favors routing workloads across cheap, frontier, and specialized APIs rather than standardizing on one model.

Originally reported by morphllm.comRead the source →
Related coverage
All OpenAI news →