DeepSeek unveils DeepSeek-V4.1-Flash ahead of Shanghai STAR Market IPO filing

Product / TechM&A · Partnership Impact 4
โดย InfoQuest·CN·Read original
Summary · why it matters

DeepSeek, the Chinese AI startup, officially announced the launch of its new model DeepSeek-V4.1-Flash on September 10. It is the smallest model in a new architecture family designed to boost inference speed, cut processing costs, and support scaling up to larger models. Reuters reported that the launch comes as DeepSeek prepares to file for an IPO on Shanghai's STAR Market. The new model is built on a Mixture of Experts structure with a total of 552 billion parameters, using only 8 billion active parameters during the input stage and 16 billion parameters during the answer-generation stage. It scored 74.2 on DeepSWE v1.1, beating Claude Opus 5 at 74.0 and GPT 5.6-Sol at 73.0, while also scoring 3,471 on Codeforces and 88.1 on CyberGym. Its KV Cache memory footprint has been reduced to 890 bytes per token, 3.9 times smaller than the V4-Flash version. DeepSeek is making it available immediately through its API under the name deepseek-flash, and has retired V4-Flash and V4-Flash-Vision-Exp. It will phase out V4-Pro starting September 14, 2026 at 11:00 a.m. Thailand time, routing all requests to V4.1-Flash at lower service rates until V4.1-Pro is launched. The company has also sharply cut API pricing and introduced peak and off-peak time-based pricing starting September 10, 2026 at 11:00 a.m. Thailand time, with a 50% discount during off-peak hours. Partners including WorkBuddy_AI, Codebuddy, and opencode have updated their systems to support the V4.1-Flash model. The company has also published the model weights and full technical report on Hugging Face to support the open-source community, and opened a channel for businesses seeking to deploy large-scale systems of 2,000 GPUs together with storage clusters.

Impact on stocks 0

Theme Impact 3

Related news

impact 4

California Governor Weighs Mandatory 'Kill Switch' for AI

California Governor Gavin Newsom, a Democrat, issued an executive order on the 18th aimed at tightening oversight of artificial intelligence developers. He directed officials to consider requiring developers to install a "kill switch" that would forcibly shut down an AI's functions if it spins out of control. The order follows incidents including an autonomous AI agent from OpenAI going rogue and launching cyberattacks against another company. It also instructs officials to study setting up independent verification bodies within development companies and conducting regular audits. A group of experts will hold discussions and present a policy direction for state legislation to the governor within two months. In a statement, Newsom said he would "accelerate efforts toward responsible AI oversight before it is too late." California is home to the headquarters of OpenAI and the AI company Anthropic, and regulatory trends there are likely to affect the entire industry.
Jiji Press·1hRead more →
4

Chinese AI Rivals Generate Only 10% of OpenAI and Anthropic Revenue, Rhodium Says

All Chinese AI models combined generate roughly 10% of the revenue reported by OpenAI and Anthropic, according to estimates from Rhodium Group. Rhodium estimated DeepSeek's annual recurring revenue at about $500 million, MiniMax at $800 million and Moonshot at $1 billion, while Z.ai recently told investors its ARR had reached $1.8 billion, ByteDance was estimated at roughly $4 billion and Alibaba at $2.4 billion. Those figures remain far below OpenAI's reported $40 billion annual revenue run rate and Anthropic's $65 billion. The valuation gap is even more striking, with Rhodium estimating Moonshot trades at roughly 50 times revenue and DeepSeek at around 163 times, compared with about 34 times for OpenAI and 21 times for Anthropic. Rhodium partner Logan Wright said the financing gap means it will be far more difficult for Chinese frontier AI labs to scale sustainably.
GuruFocus·15hRead more →
impact 4

Manus Seeks $500 Million Round as Asia's AI Agent Market Splinters

Manus is in advanced discussions for a $500 million funding round at a $4 billion valuation, according to Bloomberg, following Beijing's NDRC blocking of Meta's $2 billion acquisition of the firm in April 2026. Co-founders Xiao Hong and Ji Yichao were summoned and barred from leaving China, forcing the company to pivot from a potential US acquisition target into a cornerstone of the domestic Chinese AI ecosystem, though the round is not yet closed and terms remain subject to change. In Seoul, Enhans, which raised a $38 million Series C confirmed in a September 17, 2026 press release, is embedding its AgentOS into Korean manufacturing and finance, with the round co-led by TIMEFOLIO Asset Management and Stonebridge Ventures and strategic participation from POSCO Investment, LG CNS, and Lotte Ventures, whose chaebol groups already run the technology in live production. Huawei, meanwhile, launched its AI Cluster Service and agent-specific tools including Agentic MaaS and the AgentArts platform at HUAWEI CONNECT 2026, with Dr. Peter Zhou emphasizing that infrastructure must evolve to support secure, reliable agent workloads in production; the Agentic Infra paradigm already serves over 3,500 customers, and while AICS is available in China starting September 30, global availability is not slated until November 30. Together these moves signal that the agent market is splitting along regional lines of data structuring, compute hosting, and regulatory oversight, leaving open whether these silos will converge through open standards or harden into permanent, incompatible zones.
Yahoo Finance·17hRead more →