DeepSeek AI Releases DeepSeek-V4.1-Flash Model With 1M Context Window
The multimodal Mixture-of-Experts model features a 552B backbone and supports million-token contexts to address memory bottlenecks.
Markets on this story
Sources
Every article we clustered into this story. Headlines link to the publisher.
- DeepSeek's New Model Nearly Matches GPT-6 Astra on Design—at 1.4% of the Cost decrypt.co
- DeepSeek says new Flash AI model beats Kimi K3 on cyber, coding benchmarks South China Morning Post
- DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse marktechpost.com
- DeepSeek v4.1 Flash Hacker News
In this story
- DeepSeek AI