10 tracked signals on DeepSeek.
Import AI 465: Open vs closed gaps; Kimi K3; Demis' big policy plan
Jack Clark · Import AI · Jul 20, 2026
Open-weight models now trail closed frontier models by only 4-7 months on cybersecurity capabilities
“This implies cyber defenders have a short window to prepare before today's frontier cyber capabilities may become accessible without the same safeguards”
DeepSeek’s Insane New Architecture
Two Minute Papers · Sep 18, 2026
DeepSeek 4.1 Flash outperforms previous models and significantly reduces operational costs.
“Incredible leap forward.”
[AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale
Latent Space Blog · Sep 12, 2026
DeepSeek launches v4.1 architecture while retiring v4 Pro.
“true whalebros never wavered, and now DeepSeek are sending a weirdly mixed message by doing a completely new architecture”
DeepSeek is back... and Silicon Valley is terrified
Fireship · Aug 20, 2026
DeepSeek released a coding agent harness that is the fastest-starred GitHub repo in history
“Everything is a plugin.”
DeepSeek Just Made Closed AI Look Ridiculous
Two Minute Papers · Aug 19, 2026
DeepSeek 4 Pro achieves near-frontier quality with MIT-licensed open weights, pressuring closed AI labs.
“DeepSeek has MIT licensed open weights. Anyone can run the exact same model at their own price.”
Another DeepSeek Moment Has Arrived
Two Minute Papers · Aug 03, 2026
DeepSeek's updated flash model beats its larger pro version via post-training improvements alone
“Just the post-training step changed?”
DeepSeek’s New AI System Shouldn’t Be Possible
Two Minute Papers · Aug 26, 2026
DeepSeek released a self-modifying open-source agentic harness that extends itself on demand
“it's not you who rewrites the program, but the program rewrites itself”
[AINews] not much happened today
Latent Space Blog · Aug 01, 2026
DeepSeek V4-Flash 0731 matches GPT-5.6 performance at 60% lower cost via post-training alone
“Terminal-Bench 82.7, up +25.8 from the April preview's 56.9”
The Origins of DeepSeek
Meta (Connect) · Jun 09, 2026
DeepSeek's Liang Wenfeng built a competitive AI lab by stockpiling Nvidia A100s and reworking training to reduce GPU reliance after US export controls.
“Look, the way this works is we're going to tell you it's totally hopeless to compete with us on training foundation models. You shouldn't try, and it's your job to like try anyway. And I believe both of those things.”
DeepSeek Just Solved AI's Billion Dollar Problem
Two Minute Papers · Jun 22, 2026
DeepSeek introduces a method to fix GPU underutilization by rerouting prefill work through idle decoding machines.
“you don't need a bigger brain. You need a bigger straw.”