
AI Research
headsetAudio10 Sept 2026
DeepSeek releases V4.1 Flash with native multimodal vision, 552B MoE and lower API rates
DeepSeek has launched DeepSeek-V4.1-Flash, a 552B-parameter mixture-of-experts model featuring a novel Causal Encoder–Decoder architecture with just 8B active input and 16B active output parameters. The release brings native vision, compresses KV cache storage by up to 8x, and slashes off-peak API pricing to $0.15 per million input tokens.
smart_toyDeepSeek1.0M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$0.15/M
6 min read • 6 takeawaysReadarrow_forward