DeepSeek Statistics 2026: Users, Models, Costs and Adoption - Second Talent
Skip to content

DeepSeek Statistics 2026: Users, Models, Costs and Adoption

Matt Li By Matt Li 10 min read
TL;DR: DeepSeek had 129 million monthly app users in China in June 2026, third behind Doubao and Qwen. Outside China its chatbot is a small brand: deepseek.com drew 353.8 million visits in August, about 6% of chatgpt.com's traffic. Among developers it is one of the largest names in AI, with 371 million model downloads on Hugging Face.

Nvidia lost close to $600 billion in market value on January 27, 2025, the day DeepSeek’s app reached the top of Apple’s App Store. Coverage that week centred on one figure from the V3 report: $5.576 million, the rented GPU time for one final training run.

DeepSeek’s newest model, V4.1-Flash, reached the API on September 10, 2026 at $1.20 per million output tokens during peak hours.

Key takeaways
  1. 1At peak list prices, V4.1-Flash output costs about 17 times less than GPT-5.6 Sol and 21 times less than Claude Opus 5.
  2. 2From September 14, 2026, requests to V4-Pro run on V4.1-Flash at the Flash price.
  3. 3Eight governments or regulators have acted against DeepSeek’s app or services, five of them on government devices or agencies.
  4. 4R1, released in January 2025, is still its most downloaded model, at 48.5 million downloads.

How Many People Use DeepSeek?

QuestMobile’s 2026 half-year AI app report puts DeepSeek at 129 million monthly active users on its app in China in June 2026. Outside China the product people use is the website, and Similarweb’s estimate for deepseek.com is the closest public measure.

The app in China, June 2026
  • 129M monthly active users
  • Third among China’s AI-native apps
  • 30.0% of users in the 10-minutes-plus usage bracket
deepseek.com worldwide, August 2026
  • 353.8M visits, up 2.14% on July
  • 51.43% of desktop traffic from China
  • 8.27% from Russia, 5.86% from the US
Sources: QuestMobile 2026 half-year AI app report; Similarweb, August 2026.

ByteDance’s Doubao led China’s AI apps with 382 million monthly users and Alibaba’s Qwen had 167 million. Our Doubao vs DeepSeek comparison covers how the two apps differ.

China's top three AI-native apps by monthly active users in June 2026, drawn as circles sized by user count, from QuestMobile's 2026 half-year AI app report: Doubao by ByteDance 382 million, Qwen by Alibaba 167 million, and DeepSeek 129 million.

DeepSeek gained 2 million users between March and June. QuestMobile’s first-quarter report had it at 127 million in March, in a quarter when Qwen added 126 million. The users it has stay longer: its 30.0% share in the 10-minutes-plus bracket beat Doubao’s 27.5% and Kimi’s 26.1%.

On the web, deepseek.com is a small site next to its US rivals. Similarweb puts chatgpt.com at 5.6 billion visits in August 2026, gemini.google.com at 2.6 billion and claude.ai at 950.2 million.

Estimated total website visits in August 2026 from Similarweb: chatgpt.com 5.6 billion, gemini.google.com 2.6 billion, claude.ai 950.2 million, and deepseek.com 353.8 million, about 6 percent of chatgpt.com.

That makes DeepSeek’s site about 6% the size of ChatGPT’s and a little over a third the size of Claude’s. Our ChatGPT statistics and Claude AI statistics pages carry the figures for those two companies.

DeepSeek on OpenRouter

On OpenRouter’s model rankings, DeepSeek V4 Flash 0731 processed 12.3 trillion tokens in the seven days to September 9, 2026, the fourth highest of any model that week.

What OpenRouter measures. Only the tokens routed through OpenRouter’s own API. The figure shows developer demand on one platform, not a share of the market.

DeepSeek Models and Benchmarks

DeepSeek shipped three releases in the four weeks to September 10, 2026. The news page of DeepSeek’s API documentation lists its releases, and the timeline below picks nine.

Dec 26, 2024
V3: 671B parameters, 37B active per token
Jan 20, 2025
R1 reasoning model, five days after the app launched
May 28, 2025
R1-0528 update
Aug 21, 2025
V3.1 with hybrid thinking
Dec 1, 2025
V3.2
Apr 24, 2026
V4 preview: Pro and Flash, one-million-token context by default
Aug 13, 2026
V4-Pro general release
Aug 21, 2026
V4-Flash-Vision-Exp: the multimodal API goes live
Sep 10, 2026
V4.1-Flash: 552B parameters, reads images and text

The V4 preview in April 2026 came in two sizes: V4-Pro at 1.6 trillion parameters with 49 billion active, and V4-Flash at 284 billion with 13 billion active.

V4.1-Flash sits between them. Its model card lists these specifications:

552B parameters8B active reading, 16B writing45T training tokens1M-token contextMIT licence

The gap between reading and writing comes from a new design. The model card describes a 40-layer transformer split into a 20-layer encoder and a 20-layer decoder. DeepSeek’s release note says its key-value cache needs a quarter of the GPU memory and an eighth of the SSD storage of the previous generation. The launch took over the deepseek.com homepage on release day.

The deepseek.com homepage on September 10, 2026, announcing the release of DeepSeek-V4.1-Flash, with a chat box offering DeepThink and Search and buttons for Chat with DeepSeek and the API platform.

On DeepSeek’s own tests, V4.1-Flash resolves 74.2% of DeepSWE tasks, which ask a model to fix real software issues. Claude Opus 5 scores 74.0% and GPT-5.6 Sol 73.0% in the same table. On Terminal-Bench 4.0, a harder agent test, V4.1-Flash scores 31.2% against 51.8% for Opus 5.

Benchmark scores reported by DeepSeek on the V4.1-Flash model card at maximum reasoning effort. DeepSWE v1.1 percent resolved: Claude Opus 5 74.0, GPT-5.6 Sol 73.0, DeepSeek V4-Pro 62.7, DeepSeek V4.1-Flash 74.2. Terminal-Bench 4.0 pass at 1: Claude Opus 5 51.8, GPT-5.6 Sol 39.9, DeepSeek V4-Pro 12.4, DeepSeek V4.1-Flash 31.2.

DeepSeek ran all of these tests itself, so treat the scores as vendor claims. Our DeepSeek for coding review looks at how the models handle development work, and our DeepSeek model guide walks through the lineup one model at a time.

How Much Does the DeepSeek API Cost?

DeepSeek V4.1-Flash costs $0.30 per million input tokens and $1.20 per million output tokens at peak hours on DeepSeek’s pricing page, with off-peak rates at half that. Peak hours run 01:00 to 04:00 and 06:00 to 10:00 UTC on weekdays. The new prices took effect at 04:00 UTC on September 10, 2026.

API list prices per million tokens at short-context rates, read from vendor pricing pages on September 10, 2026, input on the left and output on the right: DeepSeek V4.1-Flash off-peak 0.15 and 0.60 dollars, DeepSeek V4.1-Flash peak 0.30 and 1.20 dollars, GPT-5.6 Luna 0.20 and 1.20 dollars, Claude Haiku 4.5 1 and 5 dollars, Claude Sonnet 5 2 and 10 dollars, GPT-5.6 Sol 4 and 20 dollars, and Claude Opus 5 5 and 25 dollars.

OpenAI lists GPT-5.6 Sol at $20 per million output tokens, a promotional price it says runs to at least November 21, 2026. Anthropic lists Claude Opus 5 at $25. By our arithmetic, DeepSeek’s peak output price is about 17 times below Sol and 21 times below Opus 5. Input runs in the same order: $4 per million tokens for Sol and $5 for Opus 5, against DeepSeek’s $0.30.

Per-token prices are not a like-for-like comparison. Anthropic says the tokenizer on Claude 4.7 and later models produces about 30% more tokens for the same text.

Cached input matters more for agents than the headline rate. DeepSeek says cache-hit charges often make up a large share of agent costs, and it charges $0.006 per million cached input tokens at peak.

DeepSeek API, per million tokensV4.1-Flash (deepseek-flash)V4-Pro
Input, cache hit (peak / off-peak)$0.006 / $0.003$0.044 / $0.022
Input, cache miss (peak / off-peak)$0.30 / $0.15$1.32 / $0.66
Output (peak / off-peak)$1.20 / $0.60$3.96 / $1.98
Context length1M tokens1M tokens

V4-Pro is on its way out. DeepSeek says V4.1-Flash outperforms it, and from 04:00 UTC on September 14, 2026, requests to deepseek-v4-pro will route to V4.1-Flash at the Flash price until a V4.1-Pro ships. If your code names V4-Pro, your bill drops and your model changes on that date.

V4-Flash and V4-Flash-Vision-Exp are already retired. Their model names, deepseek-v4-flash and deepseek-v4-flash-vision-exp, route to V4.1-Flash for compatibility.

How Much Did DeepSeek Cost to Train?

DeepSeek has not published a total cost for building its models. Its two public figures cover training runs: $5.576 million for V3’s final run and $294,000 for R1’s reasoning training, both priced at an assumed $2 per GPU hour.

The DeepSeek-V3 technical report counts 2.788 million H800 GPU hours on a 2,048-GPU cluster. In its own words, the figure “include[s] only the official training of DeepSeek-V3, excluding the costs associated with prior research and ablation experiments.”

What DeepSeek's training-cost figures cover: 5.576 million dollars for V3's final training run, 2.788 million H800 GPU hours at an assumed 2 dollars an hour excluding prior research and ablations; 294,000 dollars for R1's reasoning training, 147,000 GPU hours on top of V3; a 2,048 H800 GPU cluster trained on 14.8 trillion tokens; and hardware, salaries, data and failed runs, which neither figure includes.

That makes $5.576 million the rental value of one run. It leaves out the GPUs DeepSeek owns, its salaries, its data and the experiments that did not ship.

R1’s $294,000 comes from the supplementary information to the peer-reviewed R1 paper in Nature, published in September 2025. It covers 147,000 H800 GPU hours for R1-Zero, the supervised data creation and R1 itself, on top of V3.

DeepSeek Company Facts and Funding

Liang Wenfeng founded DeepSeek, which operates through Chinese entities including Hangzhou DeepSeek Artificial Intelligence and Beijing DeepSeek Artificial Intelligence, the two companies named in Italy’s 2025 privacy order.

Its first outside funding round surfaced in June 2026. Reuters reported that DeepSeek was set to raise about 50 billion yuan ($7.4 billion) at a post-money valuation of $52 billion to $59 billion. Its sources said Liang committed 20 billion yuan of his own money, with Tencent considering 10 billion yuan and battery maker CATL 5 billion yuan.

Those figures come from people with knowledge of the deal. DeepSeek did not respond to Reuters and has not confirmed them.

Where Is DeepSeek Banned or Restricted?

Most DeepSeek restrictions cover government devices. Italy, South Korea and Berlin went further and acted on the public app.

Government devices and systemsJan 31 to Jul 10, 2025
5 actions
Taiwan, Texas, Australia, New York and the Czech Republic
The public appJan 30 to Jun 27, 2025
3 actions
Italy, South Korea and Berlin
Our count of the notices in the table below.

Each entry links to the regulator’s or government’s own notice, apart from South Korea, where the dates come from news reports quoting its privacy commission.

JurisdictionDateWhat it restricts
ItalyJanuary 30, 2025Garante order limiting DeepSeek’s processing of Italian users’ data, with immediate effect
TaiwanJanuary 31, 2025Ministry of Digital Affairs bars public agencies from using DeepSeek AI services
Texas, USJanuary 31, 2025Governor’s ban on state employees and contractors using it on work devices
AustraliaFebruary 4, 2025PSPF Direction 001-2025 removes DeepSeek from all federal government systems and devices
New York, USFebruary 10, 2025Statewide ban on state-managed government devices and networks
South KoreaFebruary 2025 to April 28, 2025New downloads suspended at the privacy commission’s request, reported February 17; service resumed April 28
Berlin, GermanyJune 27, 2025Data protection authority reports the app to Apple and Google as unlawful content under the Digital Services Act
Czech RepublicJuly 10, 2025NÚKIB warning covering DeepSeek products, apps and APIs, alongside a government ban in state administration

Most of these notices name DeepSeek’s apps, websites and APIs, which send data to DeepSeek’s servers. Self-hosted open weights keep data on your own hardware. Whether a rule reaches self-hosted weights depends on its wording. Texas’s ban also covers state contractors on work devices, so a company selling to the state has to check it.

DeepSeek Open-Weight Downloads

Developers have downloaded DeepSeek’s models 371 million times on Hugging Face, by our sum from the Hub API across the public models on its Hugging Face organisation page as of September 10, 2026.

171M
Downloads of the 10 models in the R1 family
105
Public models on DeepSeek’s Hugging Face page
145,152
Followers, ahead of Qwen (103,651) and Meta (85,718)
Source: Hugging Face Hub API, read September 10, 2026.

Six of the ten R1-family models are smaller distilled versions built on Qwen and Llama. DeepSeek-OCR, a document-reading model, sits second on the list behind R1 itself.

Treemap of all-time Hugging Face downloads of DeepSeek's 105 public models as of September 10, 2026, 371 million in total: DeepSeek-R1 48.5 million, DeepSeek-OCR 34.6 million, R1-Distill-Qwen-32B 27.4 million, DeepSeek-V3.2 23.3 million, R1-Distill-Qwen-1.5B 21.0 million, DeepSeek-V3 20.5 million, DeepSeek-V4-Flash 11.6 million, DeepSeek-V4-Pro 10.4 million, and the other 97 models 174.0 million combined.
What a download is. Hugging Face counts download requests, not unique users. A company pulling one model onto many servers counts many times.

Running V4.1-Flash on your own hardware

The weights are free under the MIT licence, which allows commercial use and modification. DeepSeek sizes the job itself: its release note invites anyone “planning a large-scale deployment with 2,000 GPUs + a storage cluster” to get in touch.

The release ships without a Jinja chat template, the file most inference servers read to format prompts. DeepSeek provides a Python reference encoder instead, plus deepseek-recipe, a set of Rust libraries with Python bindings for production use. The model card recommends a one-million-token context window and a max_tokens setting of at least 256,000.

Our list of Chinese open-source LLMs covers the other open-weight models Chinese labs publish.

Build Your DeepSeek Team With Second Talent

Self-hosting the weights, wiring in DeepSeek’s own prompt encoder and moving batch work into off-peak hours are engineering jobs the API bill does not cover. We match companies doing that work with pre-vetted AI engineers in 24 hours, with employer of record across nine Asian markets. See our LLM developers, AI developers and China developers, or tell us what you are building and we will send profiles the next day.

Frequently Asked Questions

Is DeepSeek free to use?

The chat app is free. Its January 2025 app announcement launched it with no ads or in-app purchases. The API charges per token, and the open weights are free to download under the MIT licence.

Does DeepSeek publish a global user count?

No. The global DeepSeek user totals on statistics aggregator sites do not trace back to DeepSeek or a named tracker. The traceable figures are QuestMobile’s count for the app in China and Similarweb’s estimate for the website.

Is DeepSeek banned in the United States?

Not for the public, in the government sources we checked. The US restrictions we verified cover state government devices in Texas and New York.

Hire AI-native talent.

Second Talent connects companies with pre-vetted AI Talent.

Hire talent Apply as talent →
Matt Li

Written by

Matt Li is a tech-driven entrepreneur with deep expertise in global talent strategy, digital experience optimization, e-commerce, and Web3 innovation. He is the Co-Founder of Second Talent, a US-based company that connects businesses with top-tier tech professionals worldwide. Since launching the company in 2024, Matt has led its growth by leveraging technology to streamline remote hiring and scale distributed teams. With a background spanning product, operations, and innovation, Matt brings a cross-disciplinary perspective to the evolving digital economy. His work sits at the intersection of global talent, emerging technology, and scalable digital transformation.

More posts by Matt Li →
WhatsApp