Gunjo · Business Intelligence for the AI Era
← Sticker Wall JOURNEY · DETAIL

DeepSeek: A low-cost, open-source LLM dark horse hatched by quantitative private equity firm High-Flyer with GPU reserves

Founded: Liang Wenfeng · Hangzhou DeepSeek Artificial Intelligence Basic Technology Research Co., Ltd.

JOURNEY

Key Fields

FIELD STAMPS
IndustryAI / LLM
RegionChina
ScaleGiant
ChannelOther

Origin

Born in Zhanjiang, Guangdong in 1985, Liang Wenfeng earned a master's degree in Information and Electronic Engineering from Zhejiang University. In 2015, he founded the quantitative private equity firm High-Flyer Quant, starting out with high-frequency trading powered by machine learning. To enhance its quantitative strategies, High-Flyer began hoarding tens of thousands of NVIDIA GPUs (such as A100s) starting in 2019, eventually becoming known at its peak as one of the few domestic institutions with a 10,000-card cluster. Amid the LLM wave in 2023, Liang Wenfeng carved out a portion of the GPUs, talent, and capital to establish DeepSeek, positioning it for general artificial intelligence basic research with a 'three no's' starting principle: no external financing and no rush to commercialize.

Milestones

2015
Founding of High-Flyer Quant PMF
In 2015, Liang Wenfeng and Zhejiang University alumni including Xu Jin founded High-Flyer Quant (Jiuzhang Asset) in Hangzhou, entering quantitative private equity using machine learning and high-frequency trading. By 2019, its assets under management broke 10 billion RMB, and around 2021 it briefly surged past 100 billion RMB, becoming the blood-transfusion parent body for all of DeepSeek's subsequent compute and capital.
2021
Glowworm Supercomputer Card Hoarding Turning Point
High-Flyer successively invested in the Glowworm-1 and Glowworm-2 AI supercomputers, disclosing in 2021 that Glowworm-2 was equipped with about 10,000 NVIDIA A100s at a cost of roughly 1 billion RMB. At the time, this was mocked by outsiders as a private equity firm meddling in non-core businesses, but it was precisely these GPUs that gave DeepSeek an entry ticket to train large models even after the US chip controls on China in 2022.
2023
Establishment of DeepSeek Inflection Point
Liang Wenfeng officially registered Hangzhou DeepSeek, pulling engineers from High-Flyer to form a pure research team adhering to principles of taking no external VC money, avoiding the piling up of ToB projects, and not chasing short-term revenue. Initially, the company was completely blood-fed by High-Flyer's profits, with industry estimates putting High-Flyer's annual R&D funding allocation to it in the hundreds of millions of RMB.
2024
Release of DeepSeek-V2 Sparks Price War Growth
DeepSeek released its MoE architecture model V2, pressing API pricing down to 1 RMB per million input tokens—a fraction of the cost of GPT-4-class products at the time. This directly ignited collective price cuts from ByteDance's Doubao, Alibaba's Tongyi, and Tencent's Hunyuan, earning it the title of initiator of China's LLM price war, though the company's own revenue remained minimal.
2025
Release of R1 Shakes the World PMF
On January 20, 2025, DeepSeek-R1 was open-sourced, with inference capabilities matching OpenAI's o1. Official disclosures stated that the V3 base model training cost only about $5.576 million using 2,048 H800 GPUs for about two months. On January 27, DeepSeek climbed to #1 on the US AppStore free charts, surpassing ChatGPT. That day, NVIDIA's market capitalization evaporated by approximately $589 billion, setting a record for the largest single-day stock market drop in US history.
2025
Parent Company Earnings Backlash Failure
According to reports from Huxiu and others, Liang Wenfeng channeled back a portion of High-Flyer Quant's profits to support DeepSeek's R&D. Consequently, multiple quantitative products under High-Flyer saw negative returns within 2025, placing pressure on dividends and scale. The model of using core business profits to nourish cutting-edge research showed sustainability issues for the first time, and internal debates over whether to open up to financing became public.
2026
First External Financing and IPO Rumors Turning Point
In 2026, multiple media outlets reported that DeepSeek launched its first round of financing of around 50 billion RMB, with giants like Tencent and JD.com rumored to be involved. Valuation discussions ranged between 400 billion and 480 billion RMB, alongside rumors of a sprint toward an A-share IPO. Liang Wenfeng, who had long insisted on zero financing, was reported to have personally shelled out around 20 billion RMB to maintain controlling stake, marking the greatest posture shift in its development history from zero financing to a hundred-billion-dollar IPO.

Turning Points

  • In 2021, enduring mockery, High-Flyer hoarded about 10,000 A100s, accidentally becoming DeepSeek's ultimate trump card to bypass chip controls later.
  • In January 2025, R1 was open-sourced and topped the US AppStore, erasing approximately $589 billion from NVIDIA's market cap in a single day.
  • In 2026, shifting from zero financing to accepting roughly 50 billion RMB in external funding and floating IPO rumors signaled the transition of a pure research lab model yielding to capitalization.

Failures & Pitfalls

  • Early on, High-Flyer was mocked in the industry for private equity buying GPUs and neglecting its proper business, with the Glowworm supercomputer investing around 1 billion RMB with no commercial return in sight for years.
  • In 2025, to blood-feed R&D, Liang Wenfeng clawed back High-Flyer profits, causing multiple quantitative products under its umbrella to post negative annual returns and damaging the core business's profit-generating capacity.
  • The price war initiated by V2 drove API prices down to 1 RMB per million tokens; after the industry collectively cut prices, DeepSeek itself had virtually no profit-generating revenue.
  • After skyrocketing in popularity, model services were frequently overwhelmed and API top-ups were temporarily suspended, exposing that inference compute reserves severely lagged behind user growth.

关键成功要素

  • Relying on High-Flyer Quant's profits and a cluster of around 10,000 A100s, starting without depending on external capital or owing investors growth commitments.
  • Using engineering innovations like MoE architecture and MLA attention to compress training costs to a fraction of US peers, using efficiency to hedge against chip controls.
  • Insisting on open-source weights, building brand influence on HuggingFace downloads and developer word-of-mouth rather than advertising spend.
  • A team dominated by domestic young PhDs with flat management, where Liang Wenfeng personally participates in research to maintain technical judgment.
  • Exercising extreme restraint in financing, initially relying on the 'three no's' principle to maintain purity, and only introducing capital with a strong posture of self-funding tens of billions once its position was established.

Lessons

  • Side-project-style infrastructure investments can turn into core moats when trend winds shift; the joke about hoarding GPUs turned into the cornerstone of a hundred-billion-dollar valuation five years later.
  • A price war can instantly pierce through an industry, but the initiator must also have parent company blood-transfusion support; otherwise, low prices are a knife that hurts oneself first.
  • Open source is the fastest lever for a resource-disadvantaged party to build global prestige; hitting #1 on the AppStore once beats years of brand marketing budgets.
  • A pure research model has an expiration date; when inference compute costs rise exponentially, even the proudest team must open its mouth to the capital markets.
  • Maintaining control is more important than taking money; Liang Wenfeng preferred to fund 20 billion RMB out of his own pocket rather than let external capital dominate the board of directors.

Core Data

  • Base model training cost:Approximately $5.576 million (public data basis, independent verification unverified)
  • Training GPU configuration:2,048 H800s for about two months (public data basis, independent verification unverified)
  • Glowworm-2 GPU count:Approximately 10,000 units (public data basis, independent verification unverified)
  • 2026 first-round financing scale:Approximately 50 billion RMB (public data basis, independent verification unverified)
  • Valuation discussion range:Approximately 400 billion to 480 billion RMB (public data basis, independent verification unverified)
  • Model API pricing:1 RMB per million input tokens (public data basis, independent verification unverified)
  • NVIDIA market cap vaporized on release day:Approximately $589 billion (public data basis, independent verification unverified)

Competitors / Peers

Domestically benchmarking against Alibaba Tongyi Qianwen, ByteDance Doubao, Moonshot AI's Kimi, and Zhipu AI: Tongyi follows an open-source family-bucket plus cloud monetization route; Doubao relies on ByteDance traffic for C-end applications; Kimi focuses on long-text and secured major investment from Alibaba; Zhipu pursues a government-and-enterprise ToG route and has rushed toward an IPO. Internationally benchmarking against OpenAI and Anthropic, both of which are sprinting ahead with valuations in the hundreds of billions and annual revenue targets in the tens of billions of dollars, DeepSeek has torn open a gap beneath their pricing systems with open-source and ultra-low costs, though its commercial revenue remains far smaller than all tier-one competitors.