Scmp iconScmpSep 10, 2026

DeepSeek says new Flash AI model beats Kimi K3 on cyber, coding benchmarks

DeepSeek has released its V4.1 Flash model, claiming it outperforms its previous flagship while cutting inference costs and boosting speeds – the latest salvo in China's aggressive price-and-performance war. While built on a massive 552 billion-parameter framework, it relies on a Mixture-of-Experts (MoE) design.

DeepSeek says new Flash AI model beats Kimi K3 on cyber, coding benchmarks

Share this story

Send the public story page.

Useful takeaways from this story.

DeepSeek has released its V4.1 Flash model, claiming it outperforms its previous flagship while cutting inference costs and boosting speeds – the latest salvo in China's aggressive price-and-performance war.

While built on a massive 552 billion-parameter framework, it relies on a Mixture-of-Experts (MoE) design.

Thursday that V4.1 Flash used a new "Causal-Encoder-Decoder" architecture.

Building the complete brief

The page is ready to read now. The fuller skim-friendly version will appear here automatically.

The useful part

DeepSeek has released its V4.1 Flash model, claiming it outperforms its previous flagship while cutting inference costs and boosting speeds – the latest salvo in China's aggressive price-and-performance war. While built on a massive 552 billion-parameter framework, it relies on a Mixture-of-Experts (MoE) design. Thursday that V4.1 Flash used a new "Causal-Encoder-Decoder" architecture.

Details worth keeping

Thursday that V4.1 Flash used a new "Causal-Encoder-Decoder" architecture. In traditional AI, every query runs through the entire...

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app