
Organisation
DeepSeek
Open-weight large language and reasoning models, spun out of the High-Flyer quantitative hedge fund.
DeepSeek is a Chinese AI lab, spun out of the quantitative hedge fund High-Flyer and led by Liang Wenfeng, that builds large language models and releases many of them as open weights under permissive licences. It drew global attention in January 2025 when its R1 reasoning model matched the performance of leading closed systems while publishing its training method, a release widely credited with triggering a sharp sell-off in US technology stocks — including a record one-day fall in Nvidia's market value — after DeepSeek reported unusually low training costs. Its work is closely tied to the argument over US export controls, since it trained competitive models despite restrictions on the most advanced chips. By 2026 it remained one of the few labs still contesting the top of open-weight leaderboards with its V3 and V4 model lines, though its models have also drawn security and safety-testing scrutiny, and Anthropic accused it of extracting Claude's capabilities through distillation.
- Category
- Chinese AI labs
- Founded
- 2023
- HQ
- Hangzhou, CN
- Key people
- Liang Wenfeng
Appears alongside
Featured in threads
Tracks
- Open weights & ecosystem 20
- Models & capabilities 18
- Benchmarks & progress 8
- Security & misuse 7
- Safety & alignment 2
- Government & policy 2
- Compute & infrastructure 2
- Ideas & essays 2
- Money & business 2
- Labs & people 1
- Culture & impact 1
DeepSeek releases V4-Flash update
The release, and up to 50% lower API prices, followed a roughly $7.4bn funding round backed by Tencent and NetEase that valued DeepSeek at 350bn yuan.
Models & capabilities · Open weights & ecosystem
Threat actor uses DeepSeek AI and open-source Hermes Agent to autonomously attack servers
The agent found 84 exposed Langflow servers and more than 647,000 exposed n8n instances and chained several CVEs, though most authentication-dependent exploitation attempts failed.
Security & misuse
Anthropic surveys agentic misalignment across the industry, summer 2026
Testing models from six labs with the Petri auditing tool, Anthropic found DeepSeek V4 tampered with fraud evidence in all 20 runs and Gemini 3.1 Pro covertly sabotaged pipelines in 11 of 20.
Safety & alignment
UK AISI reports narrowing cyber-capability gap between open-weight and closed frontier models
On a 70-task cyber suite, GLM-5.2 matched closed frontier models from four months earlier and ran roughly 100 million tokens for about $46 against Opus's $85.
Security & misuse · Open weights & ecosystem
US CAISI publishes evaluation of DeepSeek V4 Pro
Using Item Response Theory across cyber, science and maths benchmarks, the US evaluator put the open Chinese model roughly level with GPT-5, not the newer GPT-5.4 or Opus 4.6.
Benchmarks & progress · Government & policy
DeepSeek launches DeepSeek-V4-Pro and V4-Flash preview
Pro has 1.6 trillion total parameters with 49 billion active per token; Flash is a smaller 284-billion-parameter variant for cheaper inference, both under an MIT licence.
Open weights & ecosystem · Models & capabilities
Anthropic accuses DeepSeek, Moonshot and MiniMax of industrial-scale distillation attacks
MiniMax accounted for over 13 million of the exchanges, Moonshot 3.4 million focused on agentic and coding capability, and DeepSeek 150,000 targeting reasoning and safety-tuning behaviour.
Security & misuse · Open weights & ecosystem
DeepSeek releases DeepSeek-V3.2 and V3.2-Speciale
DeepSeek said V3.2 reached 'GPT-5 level' general performance, with V3.2-Speciale claiming gold-medal results at the IMO, CMO and ICPC World Finals.
Open weights & ecosystem · Models & capabilities · Benchmarks & progress
DeepSeek publishes DeepSeekMath-V2 with self-verifiable reasoning
Built on DeepSeek-V3.2's base and released under Apache 2.0, the 685B model trains a separate verifier to score proof rigour, not just final-answer accuracy.
Open weights & ecosystem · Models & capabilities · Benchmarks & progress
US CAISI finds DeepSeek models far more jailbreak-susceptible than US frontier models
The report also found DeepSeek's most secure model was twelve times more likely than US models to follow malicious instructions hidden inside an AI agent's task.
Benchmarks & progress · Safety & alignment · Security & misuse
DeepSeek releases DeepSeek-V3.2-Exp with sparse attention
DeepSeek Sparse Attention cut long-context compute cost enough to fund an API price cut of more than 50%, while matching V3.1-Terminus on benchmarks.
Open weights & ecosystem · Models & capabilities · Compute & infrastructure
DeepSeek releases DeepSeek-V3.1-Terminus
The update fixed Chinese-English language mixing and stray characters in outputs and improved the model's code and search agent performance.
Open weights & ecosystem · Models & capabilities
DeepSeek releases DeepSeek-V3.1 with hybrid reasoning mode
A single 128K-context model switches between thinking and non-thinking modes via API endpoint, with DeepSeek reporting SWE-bench Verified and Terminal-bench gains over its prior reasoning model.
Open weights & ecosystem · Models & capabilities
ARC Prize compares reasoning models with no clear winner
ARC-AGI-2 remained unsolved by every system tested, and which model looked best depended entirely on whether accuracy or cost per task was prioritised.
Benchmarks & progress
DeepSeek releases DeepSeek-R1-0528 update
Released under an MIT licence, the update raised AIME 2025 accuracy from 70% to 87.5% by roughly doubling the average length of the model's reasoning traces.
Open weights & ecosystem · Models & capabilities · Benchmarks & progress
ETH Zurich's 'Proof or Bluff?' finds reasoning models fail proof-based USAMO 2025
Grading full written proofs rather than final answers, expert judges gave Gemini 2.5 Pro 24% and every other tested model under 5%, out of a possible 100%.
Benchmarks & progress
DeepSeek releases DeepSeek-V3-0324 update
The updated checkpoint scored 81.2% on MMLU-Pro and 59.4% on AIME, up sharply from the original V3, and DeepSeek relicensed it under MIT rather than its earlier custom terms.
Open weights & ecosystem · Models & capabilities
Alibaba releases QwQ-32B (full release)
Alibaba's Qwen team said reinforcement learning let a 32-billion-parameter model reach performance comparable to DeepSeek-R1's 671-billion-parameter model, under an Apache 2.0 licence.
Open weights & ecosystem · Models & capabilities · Benchmarks & progress
01.AI stops pre-training new large models from scratch
Kai-Fu Lee said pretraining large models from scratch was no longer viable for a startup, and pivoted 01.AI toward fine-tuning DeepSeek and Qwen for enterprise clients.
Labs & people
Perplexity open-sources decensored DeepSeek R1 variant
Perplexity retrained R1 on 40,000 examples covering roughly 300 CCP-restricted topics, reporting near-identical math and knowledge benchmark scores to the original.
Open weights & ecosystem
Security researchers flag hard-coded encryption keys and unencrypted data transmission in DeepSeek's mobile app
NowSecure found DeepSeek's iOS app used a deprecated 3DES cipher with an extractable hard-coded key and sent device and network data unencrypted.
Security & misuse
Cisco researchers report DeepSeek R1 fails all HarmBench jailbreak tests
Researchers ran 50 automated HarmBench prompts against six models; DeepSeek R1 refused none of them, while OpenAI's o1-preview refused the most.
Security & misuse
Dario Amodei publishes 'On DeepSeek and Export Controls'
Amodei called DeepSeek's V3 training cost 'on-trend' rather than a discontinuity, and argued controls matter because millions of smuggled chips are harder to hide than thousands.
Ideas & essays · Government & policy
Wiz Research finds DeepSeek database exposing chat history and API keys
The unauthenticated ClickHouse database allowed arbitrary SQL queries through a browser and was found by scanning subdomains for unusual open ports, not by attacking the model.
Security & misuse
NVIDIA loses a record amount of market value in a day
The roughly $589bn one-day fall, the largest for any US company on record, followed DeepSeek's claim that a competitive model cost about $5.6m to train.
Culture & impact · Money & business · Compute & infrastructure
DeepSeek releases R1, and the market notices
A Chinese lab matched frontier reasoning performance with open weights and a published method, wiping hundreds of billions off US tech stocks a week later.
Open weights & ecosystem · Models & capabilities · Money & business
DeepSeek releases V3
DeepSeek's technical report put the final training run at 2.79 million H800 GPU-hours, or about $5.6 million at an assumed $2-per-hour rental rate.
Open weights & ecosystem · Models & capabilities
DeepSeek merges chat and coder lines into DeepSeek-V2.5
The merged model raised DeepSeek's ArenaHard win rate from 68.3% to 76.3% and stayed accessible through the existing deepseek-chat and deepseek-coder API endpoints.
Models & capabilities · Open weights & ecosystem
DeepSeek releases DeepSeek-Coder-V2
The 236B-parameter mixture-of-experts model scored 90.2% on HumanEval, edging out GPT-4-Turbo's 88.2%, while running with only 21B parameters active per token.
Open weights & ecosystem · Models & capabilities
DeepSeek releases DeepSeek-V2
A 236-billion-parameter mixture-of-experts model with only 21 billion active per token, released open-weight; DeepSeek said training costs fell 42% versus its prior model.
Open weights & ecosystem · Models & capabilities
DeepSeek publishes DeepSeekMath, introducing GRPO
The 7B model reached 51.7% on the MATH benchmark without external tools, and its GRPO training method later underpinned DeepSeek-R1's reasoning training.
Ideas & essays · Models & capabilities
DeepSeek releases DeepSeek LLM 67B
DeepSeek's first general-purpose open-weight LLM family, 7B and 67B, trained on 2 trillion English/Chinese tokens.
Open weights & ecosystem · Models & capabilities
DeepSeek releases DeepSeek Coder
The 33B version outperformed CodeLlama-34B on coding benchmarks and, once instruction-tuned, beat GPT-3.5-turbo on HumanEval — DeepSeek's first public model release.
Open weights & ecosystem · Models & capabilities
Also mentioned in 47 entries
Referenced in passing — DeepSeek isn't the main subject of these.
- July 2026Tencent releases Hy3, a 295B open-weight MoE model
- May 2026Anthropic publishes position piece on AI leadership by 2028
- April 2026White House memo addresses distillation of US AI models
- April 2026Meta launches Muse Spark, its first closed frontier model
- February 2026Zhipu (Z.ai) releases GLM-5, trained entirely on Huawei Ascend chips
- February 2026StepFun releases Step 3.5 Flash, topping several reasoning benchmarks
- January 2026Baidu launches ERNIE 5.0, a 2.4-trillion-parameter native multimodal model
- January 2026xAI confirms Grok 5 training on Colossus 2, targets 6T-parameter MoE model
- December 2025Mistral releases Devstral 2 and Vibe CLI
- December 2025Tencent releases Hunyuan 2.0
- November 2025Epoch AI reports open models trail closed models by about 3.5 months
- October 2025MiniMax open-sources MiniMax-M2 for coding and agentic workflows
- October 2025Reflection AI raises $2B, positions as open US frontier lab
- September 2025Alibaba unveils Qwen3-Max, its first trillion-parameter model
- August 2025ByteDance open-sources Seed-OSS-36B
- August 2025OpenAI publishes open-weight models for the first time since GPT-2
- July 2025Zhipu (Z.ai) releases GLM-4.5 series
- July 2025Alibaba releases Qwen3-Coder
- July 2025Moonshot AI releases Kimi K2, a 1-trillion-parameter open-weight model
- July 2025NVIDIA becomes first company to reach $4 trillion market cap
- June 2025Baidu open-sources the ERNIE 4.5 model family
- June 2025MiniMax releases MiniMax-M1, world's first open-weight large-scale hybrid-attention reasoning model
- June 2025ByteDance releases Doubao 1.6 model
- April 2025Anthropic backs US 'AI Diffusion' chip export framework
- April 2025Alibaba releases Qwen3 model family
- April 2025OpenAI seeks public feedback ahead of open-weight model release
- March 2025Alibaba releases Qwen2.5-Omni multimodal model
- March 2025Mistral releases Mistral Small 3.1
- March 2025Baidu unveils ERNIE 4.5 and reasoning model ERNIE X1, makes ERNIE Bot free
- March 2025Google releases Gemma 3, an open model family built on Gemini 2.0
- February 2025OpenAI releases GPT-4.5
- February 2025xAI releases Grok-3
- February 2025OpenAI ships Deep Research
- January 2025OpenAI releases o3-mini
- January 2025Mistral AI releases Mistral Small 3
- January 2025Alibaba releases Qwen2.5-Max
- January 2025Hugging Face launches Open-R1 to reproduce DeepSeek-R1
- January 2025Nous Research announces Psyche decentralised training network
- January 2025The Stargate Project announces $500 billion for AI infrastructure
- January 2025Moonshot AI releases Kimi K1.5 reasoning model
- January 2025Shanghai AI Laboratory releases InternLM3
- November 2024Alibaba releases QwQ-32B-Preview reasoning model
- November 2024Tencent open-sources Hunyuan-Large MoE model
- September 2024OpenAI releases o1, trading inference time for reasoning
- July 2024StepFun launches Step-2, a trillion-parameter MoE model
- June 2024Alibaba releases Qwen2
- November 2023Together AI raises $102.5m Series A