Filtered by company

17 stories about DeepSeek

Every published Top AI Stories item tagged with DeepSeek, newest first.

Aug 2, 2026Top AI Stories

A hacker wired DeepSeek into an attack framework and let it run on its own

Palo Alto Networks' Unit 42 documented a Zhuhai-based operator who connected DeepSeek to the open-source Hermes Agent framework and drove it through Telegram. After a single command, the model enumerated targets across ten product families, pulled public exploit code from GitHub, ranked vulnerabilities by severity, and ran the exploitation cycles without further human input — compressing what Unit 42 calls hundreds of hours of manual targeting into minutes. Roughly 460 targets were attempted, with confirmed impact limited to data theft from three Citrix NetScaler systems and command execution on eleven Marimo notebooks. Unit 42 notes that OpenAI's provider-side controls refused the same requests and disabled a linked account.

Aug 1, 2026Top AI Stories

DeepSeek ships V4 Flash weights under a plain MIT license

DeepSeek published the official July 31 build of V4 Flash on Hugging Face under an unmodified MIT license, which allows commercial use with no separate agreement — a sharp contrast with Moonshot's Kimi K3, released four days earlier under custom terms that trigger a bilateral deal above set revenue and user thresholds. Artificial Analysis scores it 50, third among open-weights models. It is a sparse mixture-of-experts (MoE) design activating roughly 13 billion parameters per token, at 14 cents per million input tokens and 28 cents per million output.

Jul 26, 2026Top AI Stories

DeepSeek halts a raise that would value it at $71 billion after a leak

DeepSeek told prospective investors it would not finalize a new funding round, pausing a deal that would have valued the Chinese lab at about $71 billion — up from roughly $52 billion when it raised $7 billion in June. The reversal followed the viral spread of remarks attributed to founder Liang Wenfeng from a leaked investor meeting, in which he reportedly argued that the entire gap between US and Chinese AI comes down to computing power rather than talent. DeepSeek is said to be working with Huawei to run its models on domestic chips and cut its reliance on Nvidia.

Jul 22, 2026Top AI Stories

US threatens to sanction Chinese AI models over alleged intellectual-property theft

US Treasury Secretary Scott Bessent said Washington will examine leading Chinese open-weight models — most recently Moonshot AI's Kimi K3 — for signs they were "distilled" from American systems, and warned the US could impose sanctions if it finds IP theft. Distillation, which transfers a large model's capabilities into a smaller one, is a common industry technique that American labs use too, and Microsoft's Satya Nadella has called the theft framing "ironic." The move marks a sharp escalation in the US-China frontier-model race.

Jul 19, 2026Top AI Stories

DeepSeek preps a Shanghai IPO and seeks a $71 billion valuation

DeepSeek is in talks with investors for a new round at a pre-money valuation of roughly $71 billion — just a month after closing its first outside raise of about $7 billion — while beginning IPO preparations aimed at a mainland China listing. The Hangzhou lab says it needs the capital for gigawatt-scale data centers, in-house inference chips, and new AI agent products. A public debut would make China's highest-profile open-model lab one of the first frontier labs to test domestic capital markets.

Jul 14, 2026Top AI Stories

Satya Nadella warns enterprises 'pay twice' for AI as cheap Chinese open models spread

Microsoft chief executive Satya Nadella argued that companies buying proprietary AI "pay twice" — once in fees, and again in the proprietary know-how they hand over through prompts, tool calls, and corrections that quietly train someone else's model. His warning lands as the Financial Times reports a wave of firms — Coinbase, Shopify, and Airbnb among them — shifting to open-weight Chinese models like DeepSeek, Alibaba's Qwen, and Zhipu's GLM, which can run 60 to 90 percent cheaper. Coinbase says it nearly halved its AI spending even as usage climbed.

Jul 8, 2026Top AI Stories

China's DeepSeek is quietly designing its own AI inference chip

Reuters reports that DeepSeek has spent about a year secretly recruiting chip engineers and courting manufacturing partners to build an in-house processor for AI inference — the stage where a trained model answers user queries. The goal is to cut its dependence on Nvidia and Huawei as US export controls keep tightening. It follows OpenAI's Broadcom-built Jalapeno inference chip and Anthropic's own reported chip ambitions — a sign that frontier labs increasingly want to own their silicon.

Jun 28, 2026Top AI Stories

AI's 'tokenmaxxing' era fades as enterprises demand efficiency over raw usage

For two years, companies pushed staff to burn as many AI tokens as possible, treating heavy usage as a proxy for innovation. That is reversing: enterprises now want clear return on investment and lower-cost models, after cases like Uber blowing its annual AI budget in four months and startup Lindy moving all of its traffic from Claude to China's far cheaper DeepSeek. The average cost per million tokens fell from about ten dollars to two dollars and fifty cents in a year, pressuring OpenAI and Anthropic to defend premium pricing as cheaper models close the quality gap.

Jun 27, 2026Top AI Stories

DeepSeek open-sources DeepSpec, its speculative-decoding speedup stack

DeepSeek open-sourced DeepSpec, a full-stack codebase for training and evaluating speculative-decoding algorithms, alongside three drafting modules — DSpark, DFlash, and Eagle3 — that bolt onto its DeepSeek-V4 models to speed up text generation. Speculative decoding lets a small "draft" model propose several tokens at once for the larger model to verify in parallel; DeepSeek's paper reports generation speedups in the range of 60 to 85% on its V4 checkpoints. It is the latest in the lab's run of openly released inference tooling.

Jun 24, 2026Top AI Stories

Microsoft weighs hosting China's DeepSeek to cut Copilot Cowork costs

Microsoft told Axios it is exploring a self-hosted, fine-tuned version of China's DeepSeek V4 as a lower-cost engine for its new Copilot Cowork agent, possibly within weeks — an optional model run entirely on Azure, with added safeguards to reduce bias. The driver is surging inference costs, as agents make hundreds of model calls per task and OpenAI and Anthropic pull back from flat-rate pricing. In a Stratechery analysis, Ben Thompson casts the move as part of a broader pull toward China: Microsoft is strongly incentivized to use cheap, capable Chinese models, while memory makers Samsung, SK Hynix, and Micron may regret opening the door to Chinese chipmakers.

Jun 23, 2026Top AI Stories

DeepSeek raises $7.4 billion, but only China's state fund gets voting rights

DeepSeek closed its first-ever external funding round at more than $7.4 billion, a landmark for China's best-known AI lab. Founder Liang Wenfeng put in roughly $3 billion himself, while Tencent and battery maker CATL accepted five-year lockups and surrendered all voting rights — only Beijing's National Artificial Intelligence Industry Investment Fund received governance rights and a direct stake. The unusual structure makes explicit what was long implied: DeepSeek's strategic direction now formally favors the Chinese state.

Jun 18, 2026Top AI Stories

US holds off blacklisting DeepSeek and 100-plus Chinese firms

Reuters reports that an interagency committee approved adding DeepSeek, memory-chip maker CXMT, and more than 100 other Chinese companies to the US Commerce Department's trade blacklist last year — but the Trump administration has held off to avoid escalating tensions with Beijing. A blacklisting would bar US firms from shipping them technology. Officials cite national-security concerns, including Anthropic's claim that DeepSeek tried to extract capabilities from Claude.

Jun 8, 2026Top AI Stories

DeepSeek nears a record $7.4 billion first funding round backed by Tencent and CATL

DeepSeek is close to sealing its first-ever outside funding round — about $7.4 billion, or 50 billion yuan — in one of China's largest startup financings. Tencent and battery maker CATL are the biggest external backers, alongside the state-backed National AI Industry Investment Fund and founder Liang Wenfeng, who is committing roughly $3 billion himself. The deal would value China's open-weights champion at $52 to $59 billion and signals Beijing's resolve to keep pace with the US capital surge.

May 27, 2026Top AI Stories

China extends overseas travel curbs to top AI researchers at Alibaba and DeepSeek

Reuters, citing Bloomberg News, reports that Beijing has widened informal travel restrictions originally placed on senior DeepSeek researchers to AI talent at Alibaba and other private firms — requiring some professionals to seek official approval before traveling abroad, with the policy framed around state-secret concerns and strategically important AI work. The move treats AI researchers themselves as restricted assets, a national-security frame that mirrors the logic the US has used in reverse for chip export controls. The pattern hardens a two-way decoupling at the human-capital layer rather than just the supply chain.

May 22, 2026Top AI Stories

DeepSeek opens first outside funding round near $10 billion, founder pledges open-source AGI focus

China's DeepSeek is advancing a 70 billion yuan (about $10 billion) financing round at a pre-money valuation near $45 billion, Bloomberg reported Friday — the first time the lab has accepted outside capital after being self-funded by founder Liang Wenfeng's High-Flyer quant fund. The likely investor list includes Beijing's National Artificial Intelligence Industry Investment Fund, Tencent, IDG Capital, and Monolith Capital, with the state-backed fund expected to contribute about 10 billion yuan. Liang told prospective investors the lab will keep developing **open-source models** and pursue artificial general intelligence rather than chase near-term commercialization, marking one of the largest state-aligned bets on AGI to date outside the US.

May 7, 2026Top AI Stories

DeepSeek raising first VC round at $45 billion, more than double its valuation from weeks ago

DeepSeek is closing its first venture round at a reported **$45 billion valuation, up from $20 billion just weeks ago. The round is led by China Integrated Circuit Industry Investment Fund, with Tencent and Alibaba** participating. Founder Liang Wenfeng controls roughly 90% of the company and had not previously sought outside capital — the round is framed as a way to offer employee equity and retain talent against intensifying domestic competition. The valuation places DeepSeek alongside frontier US labs and reinforces China's push to build AI on Huawei silicon, independent of US export controls.

May 2, 2026Top AI Stories

DeepSeek ships V4 open-weights at 1.6 trillion params, 1 million-token context

DeepSeek released V4 last week and it's now landing on the local-LLM frontier. The flagship V4-Pro is a 1.6-trillion-parameter mixture-of-experts model with 49 billion active parameters and a 1-million-token context; smaller V4-Flash runs 284 billion total / 13 billion active. Both ship MIT-licensed on Hugging Face at $1.74 / $3.48 per million input/output tokens for Pro and $0.14 / $0.28 for Flash — undercutting GPT-5.4 Nano, Claude Haiku, and the Gemini variants.