DeepSeek begins 2-day beta of multimodal V4.1 Flash model via API
Chinese AI startup DeepSeek launched a limited-time beta of V4.1 Flash, an interim multimodal model with a new architecture, stronger performance and faster generation at lower cost, per TechNode and OrcaRouter. The test runs from September 8-10 via the existing API, priced at V4 Flash rates, per the company.
Bottom line — The beta's feedback form asks if V4.1 Flash can replace the premium V4 Pro, hinting at a price-tier reset.
Go deeper (6)
- DeelSeek described V4.1 Flash as a structural change, not just a post-training update, using native multimodal capabilities rather than a bolted-on vision encoder, per OrcaRouter.
- The model ID 'deepseek-v4.1-flash-expires-on-0910' signals a two-day test window with no model card, technical report, or benchmark table published, per OrcaRouter and TechNode.
- Developers measured roughly 420 output tokens per second in long generations, compared to around 128 tokens/s on the public V4 Flash endpoint, per OrcaRouter citing unofficial community tests.
- DeepSeek plans to release an official version 'very soon' after the beta, per OrcaRouter citing tipsters monitoring the API.
- The beta follows DeepSeek's April arrival of image recognition mode, marking its first multimodal capability for general users, per TechNode and DEV Community.
- Competitors in the Chinese AI space include Alibaba's Qwen3-Max-Thinking and Kimi K2.5, all offering different multimodal approaches, per DEV Community.