第一财经

DeepSeek V4.1 officially launched; the price war among models continues

原文:DeepSeek V4.1 Flash转正,模型价格战继续

In-Depth Analysis: The "Value for Money Revolution" Behind DeepSeek V4.1 Flash and the Capital Game

Hello everyone, I'm your financial journalist. Today, we're not talking about boring code, but about a "price war" and a "technological breakthrough" that are reshaping the landscape of the artificial intelligence (AI) industry.

On September 10th, DeepSeek officially released the V4.1 Flash model. If we had to describe this release in one word, it would be fast and affordable. For ordinary users, that means the model responds to your questions even faster; for developers, it means they don't get drained of their resources as quickly when using it.

But behind this is more than just a price cut. DeepSeek not only showed off its technical capabilities but also secured a strong partnership with Tencent and made some significant announcements in the capital market (including rumors about financing and an IPO).

Below, I'll break down this news into five key aspects to help you understand the logic behind this major development.

---

1. Technical Insights: How Does It Achieve Both Intelligence and Low Cost? (The Clever Use of Asymmetric Architecture)

Many non-technical readers might wonder: Generally, larger AI models (with more parameters) are smarter, so how does DeepSeek manage to be both affordable and fast?

Think of it like running a restaurant. Traditionally, whether you order a simple bowl of noodles or a fancy banquet, the entire kitchen team of 50 chefs would be involved. This ensures quality, but it's also extremely costly and inefficient.

DeepSeek's asymmetric model architecture (Causal-Encoder-Decoder) divides the kitchen into two sections:

  • Input Phase (Understanding Requests): Only an 8B (8 billion parameters) team is assigned to quickly understand your request. Since understanding instructions doesn't require complex logic, this small team is fast enough.
  • Output Phase (Generating Answers): A 16B (16 billion parameters) team is then used to carefully prepare the answer, ensuring quality.

Although the total number of parameters is 552B (552 billion), only a small portion is actually used for each response. This on-demand allocation of computing power significantly reduces unnecessary calculations.

In simple terms: Instead of everyone working overtime, only the necessary resources are used at the right time. This approach allows DeepSeek to maintain high intelligence while significantly lowering inference costs and latency.

---

2. The Cost Revolution: What Does a 50% Price Cut Mean for Developers?

The news mentions that the pricing for V4.1 Flash has been significantly reduced, especially for the "cache hit input" service, which has seen a 60% decrease.

A key concept here is the KV Cache (Key-Value Cache), which can be thought of as AI's "short-term memory." When you interact with AI repeatedly, it needs to remember previous conversations for context. This memory consumption is high on both memory (HBM) and storage (SSD).

With the new architecture, DeepSeek has reduced the memory and storage requirements by a quarter and an eighth, respectively.

What does this mean for users?

  • Faster responses: With less memory usage, AI can read and write data faster, providing a smoother experience.
  • More stable services: Servers are less stressed, making them less likely to fail during peak times.

What does this mean for developers and businesses?

  • Lower deployment costs: Previously, expensive servers were needed to run AI applications. Now, the same servers can serve more users, or the same number of users can be served with cheaper servers.
  • The start of a price war: By lowering prices, DeepSeek may drive other competitors (including proprietary models) to reduce their costs, potentially losing cost-sensitive developer customers.

In simple terms: AI, once a luxury, is now becoming more accessible. DeepSeek has reduced the underlying costs, allowing it to offer lower API usage fees.

---

3. The Ecological Alliance: Why Is Tencent Investing So Much in DeepSeek?

An interesting detail in the news is that Tencent's tools like WorkBuddy and CodeBuddy have fully integrated with V4.1 Flash, and Tencent AI publicly stated, "We also have our own models, but open-source solutions are better."

This reveals a significant shift in Tencent's AI strategy: from working independently to leveraging an ecosystem.

  • Tencent's Strategy: Tencent has a strong presence in social and business scenarios (WeChat, WeCom, Tencent Docs), but its own models may not be as cost-effective or as influential in the open-source community as DeepSeek's. Instead of developing its own model from scratch, it's better to use DeepSeek's, which offers the best value for money.
  • Capital Bond: Tencent was a major investor in DeepSeek's first round of financing, contributing about 10 billion yuan. This investment not only fosters technical cooperation but also strengthens their partnership.
  • Open Source vs Proprietary: By favoring open-source solutions, Tencent signals that open, customizable, and cost-effective models are more attractive in the enterprise market. DeepSeek's open-source nature fits well with Tencent's goal of building an open ecosystem.

In simple terms: Tencent doesn't want to reinvent the wheel; it wants to use DeepSeek's powerful technology to enhance its products.

---

4. Product Implementation: From Chatbots to "Agent Factories"

This release includes an update to the DeepSeek Harness (the AI framework) to version 0.1.5.

What is an "agent"? Traditional AI only responds to queries; modern agents can search for information, write code, operate software, and complete tasks on their own, like skilled employees.

DeepSeek Harness promotes the "everything as a plugin" philosophy, meaning:

  • Modularity: Models, tools, and skills can be combined freely like Lego pieces.
  • Optimized Integration: The V4.1 Flash model is specifically trained for the Harness, making them work together more efficiently.

Why is this important?

  • Lowered development barriers: Developers don't need to build complex AI applications from scratch; they can simply drag and drop plugins in the Harness to create functional AI components.
  • Penetration into the B2B Market: Enterprises need AI that can handle tasks like email management, data analysis, and code writing. DeepSeek is moving from the consumer (C-side) to the business (B-side) market.

In simple terms: DeepSeek is no longer just selling a "brain"; it's offering a "complete package" that enables developers to quickly build various AI applications.

---

5. The Capital Movement: 50 Billion in Financing and the Ambition for an IPO on the STAR Market

Let's look at the more sensitive capital aspects:

  • Financing Rumors: DeepSeek plans to raise 50 billion yuan in its second round of financing, with a pre-financing valuation of 500 billion yuan.
  • IPO Rumors: It has been in talks with CITIC Securities and is considering an IPO on the STAR Market, with a target listing date in 2027.

Although DeepSeek has not confirmed these rumors, they are not baseless.

Why such a high valuation?

1. Technical Advantages: The asymmetric architecture and cost optimization demonstrate its technological leadership.

2. Market Position: DeepSeek has established an ecosystem similar to Android in the open-source AI space, with high developer loyalty.

3. Strategic Value: With the country's focus on AI autonomy, DeepSeek holds significant strategic value.

What does an IPO mean?

  • Financial Support: AI is a capital-intensive industry, and DeepSeek needs funds for model training and computing power.
  • Valuation Benchmark: An IPO would set a benchmark for the Chinese AI industry.
  • Talent Attraction: The news mentions a severe shortage of engineers, and the offer of 150 new positions is a strong incentive to attract top AI talent.

Risks:

  • Profitability: The capital market will ultimately judge based on profitability. Despite low costs, DeepSeek still needs to prove it can turn a free or low-cost model into a profitable one.
  • Regulation and Competition: With competition from giants like Tencent, Alibaba, Baidu, and overseas players like OpenAI and Anthropic, DeepSeek must maintain its technological lead.

In simple terms: DeepSeek is transitioning from a tech startup to a potential tech giant. The 50-billion financing and IPO plans indicate its ambition to become a leading provider of AI infrastructure, challenging the global AI landscape.

---

Conclusion

The release of DeepSeek V4.1 Flash is a combination of technology, business, and capital:

  • Technologically, it has achieved high intelligence at low costs through the asymmetric architecture, breaking the myth that large models must be expensive.
  • Businesswise, it has strengthened its position in the cost-effective market and accelerated the adoption of AI in business applications.
  • Capitalistically, it has secured substantial funding and positioned itself for future development and competition.

For consumers, this means we'll soon enjoy cheaper, faster, and smarter AI services. For the industry, DeepSeek is using an open-source and cost-effective approach to drive the AI industry towards greater efficiency and practical applications.

"The giant has returned, and this time, it's equipped with sharper tools and more financial resources."