NVDA 214.72 ▼0.98%GOOGL 344.82 ▲1.22%MSFT 483.24 ▲0.43%AMD 473.25 ▲0.81%INTC 90.07 ▼2.24%TSMC 418.95 ▲0.71%AMZN 258.63 ▼0.57%META 549.90 ▲0.75%AAPL 309.35 ▼0.63%PLTR 179.94 ▲3.44%
Markets at last close

DeepSeek · Models

DeepSeek releases V4 Flash Vision Exp

·1 min read

DeepSeek has introduced V4 Flash Vision Exp, a multimodal addition to its flagship V4 model series. The model is available at launch through the Chinese startup’s paid developer platform, with a free release possible later because DeepSeek has open-sourced many earlier models.

V4 Flash Vision Exp is based on V4 Flash, released in April. In DeepSeek’s tests, it surpassed V4 Flash on text benchmarks except Cybergym, which evaluates software vulnerability discovery, and delivered its largest gains in image analysis. It scored more than 10% higher on two visual tests and outperformed Anthropic PBCC’s Opus 4.8 on ALE and ZeroBench, which measure complex application tasks and difficult image analysis challenges.

DeepSeek has not disclosed the new model’s architecture, but V4 Flash is described as a mixture of experts system with 284 billion parameters and component networks containing 13 billion parameters each. Its KV cache compression techniques, HCA and CSA, are said by DeepMind to reduce computing power for prompts with 1 million tokens by 73%. V4 Flash was trained on 32 trillion tokens using Muon, an algorithm designed to accelerate hidden-layer calibration.

Originally reported by siliconangle.comRead the source →
Related coverage
All DeepSeek news →