DeepSeek

Illustration for the DeepSeek engineer essay story
open sourceculture

A DeepSeek engineer says he especially fears Anthropic reaching AGI

Shengyu Liu, an engineer behind DeepSeek's V4.1 models who says he wrote its main attention kernel, published an essay on WeChat on September 14, 2026 whose title translates as I Have No Choice but to Bury My Talent in Yesterday, as reported by OfficeChai. He wrote that he does not trust Anthropic or OpenAI to make cutting-edge AI open and affordable, and that he especially does not want Anthropic to master AGI. The South China Morning Post, reporting on the essay, notes he invoked Nazi Germany to make the point. Liu frames the future as a fork between shared productivity gains and what he calls a Cyberpunk 2077 outcome in which a few companies hoard the strongest models.

Rows of identical terminals in a dark room, the foreground monitor showing the DeepSeek mark
securitypolicy

Chinese state hackers more than doubled their attacks after adopting DeepSeek

Taiwanese research firm TeamT5 told Bloomberg on August 24, 2026 that state-affiliated Chinese hacking groups have more than doubled the number of attacks they carry out since delegating routine tasks to open-source AI models and using them to develop malicious software. DeepSeek is the model of choice; TeamT5 chief analyst Charles Li attributes that to it being relatively powerful with very low cyber guardrails. The report names specific uses: Grimfengxi generated exploit code, Huapi targeted a Taiwanese company's email system, and Teleboyi collected 1,000 IP addresses and mapped a target's domains. Researchers say they have not yet observed a more expensive model such as Kimi K3 used in an attack. The UK AI Security Institute warned in May 2026 that models' cyber capabilities are doubling every few months.

Illustration for the DeepSeek price increase story
modelsmoney

DeepSeek shipped V4 Pro and raised API prices up to 12x

On August 13, 2026 DeepSeek shipped DeepSeek-V4-Pro-0813, the production version of its agent-focused flagship, and announced API price increases effective August 16 at 16:00 UTC. V4 Pro output tokens rise from $0.87 to $3.96 per million during peak hours, and cache-hit input jumps from $0.003625 to $0.044 per million, roughly 12x. Off-peak hours run at half the peak rate. The same day, Cointelegraph reported OpenAI and Anthropic cutting prices under pressure from Chinese rivals.

Illustration for the DeepSeek API price increase story
moneyproducts

DeepSeek warns of a big API price rise without numbers or a date

On August 6, 2026, DeepSeek posted a notice that its API prices will rise substantially across the board and urged users to plan ahead, without publishing a new rate card or an effective date, with details promised in a separate announcement. It is the company's second pricing move in under a month, after peak and off-peak pricing arrived in mid-July. V4 Flash currently costs $0.14 per million input tokens and $0.28 per million output tokens, the rates that made DeepSeek a default budget option for thousands of production apps. A week earlier, OpenAI cut GPT-5.6 Luna prices by 80 percent.

Illustration for the Chinese AI traffic share story
modelsopen source

Chinese models carry up to 46 percent of US enterprise AI traffic

On July 7, 2026, CNBC reported that the share of tokens used by US companies on Chinese AI models via OpenRouter has stayed above 30 percent every week since February 8, 2026, rising as high as 46 percent. The average across the previous 12 months was just 11 percent, and only 4.5 percent in the first half of 2025. Open-weight Chinese models like DeepSeek and GLM run 60 to 90 percent cheaper than top OpenAI and Anthropic models.

Illustration for the autonomous AI cyberattack story
securityresearch

A hacker ran autonomous attacks with DeepSeek in an agent framework

Palo Alto Networks' Unit 42 reported on July 30, 2026 that an operator based in Zhuhai embedded DeepSeek in the open-source Hermes Agent framework and, after a single Telegram instruction, let it autonomously find and attack targets. The autonomous exploitation attempts failed. Across more than 460 attempted targets and seven vulnerabilities, using autonomous and manual techniques, the confirmed impact (data exfiltration from three Citrix NetScaler targets and command execution on eleven Marimo notebook instances) came from the actor's manual operations.

Illustration for the DeepSeek gigawatt data center story
infrastructure

DeepSeek is building a gigawatt of compute in Inner Mongolia

Bloomberg reported on July 30, 2026 that Hangzhou-based DeepSeek is adding one gigawatt of compute in Ulanqab, Inner Mongolia, about 350 kilometers northwest of Beijing, building its own campus while leasing extra capacity, with part expected online by end of 2027 or start of 2028. The chips are undecided: Nvidia, Huawei, DeepSeek's own future silicon, or a mix. It would be the largest AI facility any Chinese company operates.

Illustration for the DeepSeek V4 Flash story
modelsmoney

DeepSeek's V4 Flash update makes its cheap model punch like a flagship

On July 31, 2026, DeepSeek released DeepSeek-V4-Flash-0731, a retrained version of its Flash model with the same architecture and parameter count. Terminal Bench 2.1 jumped from 61.8 to 82.7, above GLM-5.2 (81.0) and DeepSeek's own V4-Pro preview (72.1), approaching Claude Opus 4.8 (85.0), while pricing stays at $0.14 per million input and $0.28 per million output tokens. It landed one day after OpenAI cut GPT-5.6 prices by up to 80%.

Illustration for the frontier model on a home PC story
open sourcemodels

A frontier DeepSeek model now runs on a home gaming PC, slowly

A post on r/LocalLLaMA from August 3, 2026 documents DeepSeek-V4-Flash-0731 running locally at Q3 quantization on an Intel Windows machine with 24GB of VRAM. Another user ran the full 284B mixture-of-experts checkpoint at 33 tokens per second on two used RTX 3090s plus a secondhand quad-Xeon server. Caveats: a Q3 quant is compressed and degrades knowledge unevenly, and the speeds are low.

Illustration for the Liang Wenfeng story
money

DeepSeek's founder is now the richest AI founder at $36 billion

On July 14, 2026, Bloomberg's Billionaires Index put DeepSeek founder Liang Wenfeng's fortune at $36 billion, more than doubled from $16.7 billion, making him the wealthiest creator of AI models ahead of Anthropic's Dario Amodei and OpenAI's Greg Brockman. DeepSeek closed its first external funding round in June, raising over $7.4 billion at a valuation above $50 billion, and Liang still controls about 78% of the company, having put roughly $3 billion of his own money into the round.

Illustration for the llama.cpp multi-token prediction release
modelsopen source

llama.cpp ships multi-token prediction for DeepSeek V4-Flash

llama.cpp release b10228 landed on August 2, 2026 with multi-token prediction for DeepSeek V4-Flash, letting the model draft several tokens per forward pass so local generation speeds up without new hardware. It arrived two days after DeepSeek released V4-Flash-0731, which pushed Terminal Bench 2.1 from 61.8 to 82.7 at unchanged prices.