DeepSeek opens a two-day beta of V4.1 Flash, a new-architecture natively multimodal model

DeepSeek

Models / LLM media only 3 src. ~1 min

On September 8 DeepSeek quietly started a limited internal beta of V4.1 Flash, callable as deepseek-v4.1-flash-expires-on-0910 on the existing API endpoint with no base_url change. The company says the interim build uses a new model architecture with native multimodal input, is faster and cheaper than V4 Flash, and beta billing matches deepseek-v4-flash with a 20-concurrent-request cap; access expires automatically on September 10. A feedback survey asks testers whether it can fully replace the production DeepSeek V4 Pro.

Why it matters

First public look at DeepSeek's next-generation architecture and a signal that native multimodality is coming to the mainline Flash/Pro line.

Importance: 3/5

First glimpse of DeepSeek's next-generation architecture; 2 independent media sources

Sources