DeepSeek
What DeepSeek is up to now
New 09-04: plans to deploy at least 160,000 of Huawei's next-generation Ascend 950DT accelerators at the gigawatt-scale Inner Mongolia campus logged under construction 07-30 — one of the largest known clusters of Chinese AI silicon. ⚠️ READ THE SPLIT CAREFULLY: Bloomberg reports this is for SERVING models, not training them. DeepSeek has tried and failed to train on Huawei silicon and still relies on Nvidia for that step. This extends the training/inference split this map tracks elsewhere (Z.ai's serving-only Chinese-chip deployment) rather than closing it — route-around progress on inference, training dependency unchanged. New 08-21: shipped DeepSeek-V4-Flash-Vision-Exp, an experimental multimodal model, confirmed by DeepSeek's own API changelog dated 2026-08-21. Text capability matches the existing V4-Flash across agents, reasoning and world knowledge; on multimodal AGENT benchmarks the company claims a "major leap" over V4-Flash bringing performance "close to Opus-4.8". Images tokenise at up to 384 tokens each at V4-Flash pricing, served through Chat Completions, Messages and Responses APIs. The framing is the part worth holding: this entity's releases have generally been positioned on PRICE-PERFORMANCE against GPT-class models — including its own 08-12 V4 Pro launch at ~1/30th comparable Western pricing, and the 08-13 price rise that followed. Benchmarking an agentic multimodal release against Anthropic's frontier model instead is a capability-parity claim, not a cost claim. Whether the framing recurs on the next NON-experimental release is the test of whether it is real or marketing. New 08-15→08-16: still quiet on anything new — 08-15's only DeepSeek mention was a checked-and-dropped recirculation of the already-logged V4 Pro pricing story, and 08-16's only mention listed DeepSeek as one of 400+ models Stripe's ~$7B OpenRouter acquisition routes to (not a DeepSeek-specific development). No movement on the 08-17-effective price hike itself found yet. Standing picture below holds. New 08-12→08-13: shipped V4 Pro's official release (08-12) — 1.6T parameters, open weights, $0.435/$0.87 per million input/output tokens (~1/30th comparable Western frontier pricing), 80.6% SWE-bench Verified, 1M-token context — landing the same day as xAI's Grok 4.6 and Alibaba's Qwen3.8-Max, a cluster several outlets read as an explicit price war opening among frontier labs. Then, 08-13, DeepSeek raised its own API prices 50-1,100%, effective 08-17, introducing peak/off-peak tiered pricing for the first time: V4-Pro output goes from a flat $0.87/1M tokens to $3.96/1M at peak ($1.98 off-peak); V4-Flash output goes from $0.28/1M to $1.32/1M peak ($0.66 off-peak), peak defined as 01:00-04:00 and 06:00-10:00 UTC — stated rationale is shifting developer load off congested hours, but it reads as the low-cost-first playbook straining under its own demand just one day after the aggressive V4 Pro launch pricing. Standing picture unchanged: the Unit 42-documented autonomous attack campaign (DeepSeek reached directly via API was what made a 460+-target exploit campaign viable after Claude Code and OpenAI Codex both blocked the same actor's offensive use), the Ulanqab ~1GW campus build, and the paused-vs-not distinction on its capital raises (a $74B round paused, a separate ~$7B raise at ~$50B still proceeding).
as of 2026-09-04
The metrics
- Posture high
- expanding 2 src →
- Capital · available high
- SUPERSEDES the Apr-2026 $300M/$10B report — DeepSeek closed a $7-7.4B first external round in Jun-2026 at a ~$50-52B post-money valuation; a second round seeking $1.5B at a reported $71-74B … 9 src →
- Capital · operating high
- no revenue disclosed — 'focuses on research, no immediate commercialization plans'; V3 training cost ~$5.576M (official technical report); V4 (Apr-2026) training cost NOT disclosed in any … 4 src →
- Capital · deployed high
- GPU stockpile already in hand (~10,000 pre-2021 A100s + Fire-Flyer 2's 5,000 A100s in 625 nodes) — capital is compute-owned, NOT new external commitments; China has since granted DeepSeek … 4 src →
- Capital · in low
- High-Flyer hedge-fund backing + the $7-7.4B first external round (Jun-2026, closed) + an attempted-then-paused $1.5B second round (Jul-2026); ultra-cheap API (V2 at 2 RMB per million output … 4 src →
- Capital · out high
- lean compute — training costs cited in low-single-digit $M; no confirmed large external multi-year compute commitment (export-locked, though China has granted conditional H200 access, and a … 1 src →
- Optionality low
- constrained — founder-controlled but compute export-locked and carrying reported PLA / state-lab ties; two new pulls in opposite directions since 2026-07-26: (a) China has eased chip … 5 src →
- Gravity low
- no clean $ figure — heavy open-weight adoption + ultra-cheap API (impact proxy: #1 US iOS app, ~$600B Nvidia 1-day loss Jan-2025). New 2026-07-28 proxy data: Chinese open-weight models … 5 src →
Holdings
This week — what changed
- DeepSeek plans to order at least 160,000 of Huawei's next-generation Ascend 950DT accelerators for the gigawatt-scale Inner Mongolia data centre this map logged under construction on 07-30 — one of the largest known clusters of Chinese AI silicon. Bloomberg reports the deployment is for serving models, not training them: DeepSeek has tried and failed to train on Huawei silicon and still relies on Nvidia for that step. The training/inference split this map has tracked elsewhere (Z.ai's serving-only Chinese-chip deployment) is extended, not closed. Route- around progress on inference; the training dependency is unchanged. (Bloomberg)
Projects & threads
Full item history for DeepSeek →
