Every published Top AI Stories item tagged with DeepSeek, newest first.
Palo Alto Networks' Unit 42 documented a Zhuhai-based operator who connected DeepSeek to the open-source Hermes Agent framework and drove it through Telegram. After a single command, the model enumerated targets across ten product families, pulled public exploit code from GitHub, ranked vulnerabilities by severity, and ran the exploitation cycles without further human input — compressing what Unit 42 calls hundreds of hours of manual targeting into minutes. Roughly 460 targets were attempted, with confirmed impact limited to data theft from three Citrix NetScaler systems and command execution on eleven Marimo notebooks. Unit 42 notes that OpenAI's provider-side controls refused the same requests and disabled a linked account.
DeepSeek published the official July 31 build of V4 Flash on Hugging Face under an unmodified MIT license, which allows commercial use with no separate agreement — a sharp contrast with Moonshot's Kimi K3, released four days earlier under custom terms that trigger a bilateral deal above set revenue and user thresholds. Artificial Analysis scores it 50, third among open-weights models. It is a sparse mixture-of-experts (MoE) design activating roughly 13 billion parameters per token, at 14 cents per million input tokens and 28 cents per million output.
DeepSeek told prospective investors it would not finalize a new funding round, pausing a deal that would have valued the Chinese lab at about $71 billion — up from roughly $52 billion when it raised $7 billion in June. The reversal followed the viral spread of remarks attributed to founder Liang Wenfeng from a leaked investor meeting, in which he reportedly argued that the entire gap between US and Chinese AI comes down to computing power rather than talent. DeepSeek is said to be working with Huawei to run its models on domestic chips and cut its reliance on Nvidia.
US Treasury Secretary Scott Bessent said Washington will examine leading Chinese open-weight models — most recently Moonshot AI's Kimi K3 — for signs they were "distilled" from American systems, and warned the US could impose sanctions if it finds IP theft. Distillation, which transfers a large model's capabilities into a smaller one, is a common industry technique that American labs use too, and Microsoft's Satya Nadella has called the theft framing "ironic." The move marks a sharp escalation in the US-China frontier-model race.
DeepSeek is in talks with investors for a new round at a pre-money valuation of roughly $71 billion — just a month after closing its first outside raise of about $7 billion — while beginning IPO preparations aimed at a mainland China listing. The Hangzhou lab says it needs the capital for gigawatt-scale data centers, in-house inference chips, and new AI agent products. A public debut would make China's highest-profile open-model lab one of the first frontier labs to test domestic capital markets.
Microsoft chief executive Satya Nadella argued that companies buying proprietary AI "pay twice" — once in fees, and again in the proprietary know-how they hand over through prompts, tool calls, and corrections that quietly train someone else's model. His warning lands as the Financial Times reports a wave of firms — Coinbase, Shopify, and Airbnb among them — shifting to open-weight Chinese models like DeepSeek, Alibaba's Qwen, and Zhipu's GLM, which can run 60 to 90 percent cheaper. Coinbase says it nearly halved its AI spending even as usage climbed.
Reuters reports that DeepSeek has spent about a year secretly recruiting chip engineers and courting manufacturing partners to build an in-house processor for AI inference — the stage where a trained model answers user queries. The goal is to cut its dependence on Nvidia and Huawei as US export controls keep tightening. It follows OpenAI's Broadcom-built Jalapeno inference chip and Anthropic's own reported chip ambitions — a sign that frontier labs increasingly want to own their silicon.
For two years, companies pushed staff to burn as many AI tokens as possible, treating heavy usage as a proxy for innovation. That is reversing: enterprises now want clear return on investment and lower-cost models, after cases like Uber blowing its annual AI budget in four months and startup Lindy moving all of its traffic from Claude to China's far cheaper DeepSeek. The average cost per million tokens fell from about ten dollars to two dollars and fifty cents in a year, pressuring OpenAI and Anthropic to defend premium pricing as cheaper models close the quality gap.
DeepSeek open-sourced DeepSpec, a full-stack codebase for training and evaluating speculative-decoding algorithms, alongside three drafting modules — DSpark, DFlash, and Eagle3 — that bolt onto its DeepSeek-V4 models to speed up text generation. Speculative decoding lets a small "draft" model propose several tokens at once for the larger model to verify in parallel; DeepSeek's paper reports generation speedups in the range of 60 to 85% on its V4 checkpoints. It is the latest in the lab's run of openly released inference tooling.
Microsoft told Axios it is exploring a self-hosted, fine-tuned version of China's DeepSeek V4 as a lower-cost engine for its new Copilot Cowork agent, possibly within weeks — an optional model run entirely on Azure, with added safeguards to reduce bias. The driver is surging inference costs, as agents make hundreds of model calls per task and OpenAI and Anthropic pull back from flat-rate pricing. In a Stratechery analysis, Ben Thompson casts the move as part of a broader pull toward China: Microsoft is strongly incentivized to use cheap, capable Chinese models, while memory makers Samsung, SK Hynix, and Micron may regret opening the door to Chinese chipmakers.
DeepSeek closed its first-ever external funding round at more than $7.4 billion, a landmark for China's best-known AI lab. Founder Liang Wenfeng put in roughly $3 billion himself, while Tencent and battery maker CATL accepted five-year lockups and surrendered all voting rights — only Beijing's National Artificial Intelligence Industry Investment Fund received governance rights and a direct stake. The unusual structure makes explicit what was long implied: DeepSeek's strategic direction now formally favors the Chinese state.
Reuters reports that an interagency committee approved adding DeepSeek, memory-chip maker CXMT, and more than 100 other Chinese companies to the US Commerce Department's trade blacklist last year — but the Trump administration has held off to avoid escalating tensions with Beijing. A blacklisting would bar US firms from shipping them technology. Officials cite national-security concerns, including Anthropic's claim that DeepSeek tried to extract capabilities from Claude.
DeepSeek is close to sealing its first-ever outside funding round — about $7.4 billion, or 50 billion yuan — in one of China's largest startup financings. Tencent and battery maker CATL are the biggest external backers, alongside the state-backed National AI Industry Investment Fund and founder Liang Wenfeng, who is committing roughly $3 billion himself. The deal would value China's open-weights champion at $52 to $59 billion and signals Beijing's resolve to keep pace with the US capital surge.
Reuters, citing Bloomberg News, reports that Beijing has widened informal travel restrictions originally placed on senior DeepSeek researchers to AI talent at Alibaba and other private firms — requiring some professionals to seek official approval before traveling abroad, with the policy framed around state-secret concerns and strategically important AI work. The move treats AI researchers themselves as restricted assets, a national-security frame that mirrors the logic the US has used in reverse for chip export controls. The pattern hardens a two-way decoupling at the human-capital layer rather than just the supply chain.
China's DeepSeek is advancing a 70 billion yuan (about $10 billion) financing round at a pre-money valuation near $45 billion, Bloomberg reported Friday — the first time the lab has accepted outside capital after being self-funded by founder Liang Wenfeng's High-Flyer quant fund. The likely investor list includes Beijing's National Artificial Intelligence Industry Investment Fund, Tencent, IDG Capital, and Monolith Capital, with the state-backed fund expected to contribute about 10 billion yuan. Liang told prospective investors the lab will keep developing **open-source models** and pursue artificial general intelligence rather than chase near-term commercialization, marking one of the largest state-aligned bets on AGI to date outside the US.
DeepSeek is closing its first venture round at a reported **$45 billion valuation, up from $20 billion just weeks ago. The round is led by China Integrated Circuit Industry Investment Fund, with Tencent and Alibaba** participating. Founder Liang Wenfeng controls roughly 90% of the company and had not previously sought outside capital — the round is framed as a way to offer employee equity and retain talent against intensifying domestic competition. The valuation places DeepSeek alongside frontier US labs and reinforces China's push to build AI on Huawei silicon, independent of US export controls.
DeepSeek released V4 last week and it's now landing on the local-LLM frontier. The flagship V4-Pro is a 1.6-trillion-parameter mixture-of-experts model with 49 billion active parameters and a 1-million-token context; smaller V4-Flash runs 284 billion total / 13 billion active. Both ship MIT-licensed on Hugging Face at $1.74 / $3.48 per million input/output tokens for Pro and $0.14 / $0.28 for Flash — undercutting GPT-5.4 Nano, Claude Haiku, and the Gemini variants.