← All tags

#DeepSeek

64 stories · 35 issues · daily DeepSeek news timeline


#DeepSeek timeline · newest first
Thu, Oct 1
DeepSeek and Huawei open-source Ascend programming tools, filling a key gap for training on Chinese chips
The hard part of chip independence was never just making the chip but getting developers to write code for it, and DeepSeek is solving that toughest software problem for Huawei.
Alibaba's ModelScope and OSChina's MoArk vie to become China's Hugging Face
Open models are free; the fight is over where developers first download them, because that's often where they first pay for compute.
Wed, Sep 30
Moonshot's Kimi K3.1 surfaces on its API platform, reportedly with a 1M-token context and a launch next month
Chinese model makers are moving from chasing a benchmark score to selling one model at several price points, making pricing design as important as capability.
DeepSeek Harness v0.2 preview ships desktop installers, taking the agent tool off the command line
Competition among agent tools is shifting from whose model codes better to whose installer is easier and whose plugins are easier to find; the desktop is the new battleground.
Fri, Sep 25
DeepSeek’s annualized revenue run rate reaches $1 billion as it seeks roughly $7.5 billion
A model business must show both that customers will keep paying and that each extra dollar of revenue does not require an equivalent increase in compute spending.
Thu, Sep 24
Paper signed by Liang Wenfeng details DeepSeek's DSec agent-training system, which creates 5,000 sandboxes a second and runs 380,000 at once
The bottleneck in agent training is shifting from GPUs to CPUs and environment scheduling, and by publishing DSec, DeepSeek signals it no longer sees this layer as a moat.
Wed, Sep 23
Liang Wenfeng tells investors DeepSeek is training a 2-trillion-parameter model and eyeing 8 trillion
The lab famous for saving compute is now quoting trillions of parameters; Chinese model competition has switched its unit from efficiency to scale.
Trump tells the UN that U.S. documents will say "super intelligence," as DeepSeek briefs the Security Council
Rejecting restraint at the podium while inviting companies to discuss risk in the council — global AI governance received two opposite signals in one week.
Xiaomi's MiMo-V2.6-Pro tops the open-weight leaderboard, tied with Grok 4.7
$3 million for an open model tied with Grok 4.7 — a phone maker is repricing the entry fee to frontier capability.
Tue, Sep 22
The Information says DeepSeek aims to train on Huawei chips, with deliveries expected late this year or early next
The test of domestic compute moving from inference to frontier training is sustained cluster performance, not a shipment announcement.
OpenRouter-based estimates put Chinese models at 67.46 trillion weekly tokens, with important coverage limits
Token share shows where models are being tried and called; commercial share still depends on who pays and at what price.
Thu, Sep 17
DeepSeek engineer Liu Shengyu discusses changing roles, as elite kernel optimization begins yielding ground to AI collaboration
AI may first consume the repetitive manual optimization time of top engineers, rather than the engineering job itself.
Wed, Sep 16
DeepSeek's official desktop code is public, with Mac and Windows build targets
The confirmed desktop progress concerns build and update infrastructure; general installer availability and comparable task performance require separate verification.
Tue, Sep 15
DeepSeek lets V4.1 Flash carry a generational change, as a leak points to a 3T-parameter Code 2.0 in September
Version numbers no longer mark a capability jump; unit cost is the column that has to change.
Hillhouse VC partner Yan Wentao may become DeepSeek's first CFO, after months of jostling for the role
A company valued at RMB 500 billion filling its finance seat usually means capital moves are on a schedule.
Sat, Sep 12
DeepSeek V4.1 Flash Exposes Day-One Support Gap as Nvidia Runs vLLM Across Six GPU Families
New-model support has become an hourly contest; a usable day-one image can influence developer choice as much as hardware specifications.
DeepSeek App Adds Four Reading Voices, While Talk of a New TTS Model Remains Speculation
Four new voices are a confirmed product change; there is no evidence yet that the underlying model is new, so the feature and a model release must remain separate.
Fri, Sep 11
DeepSeek Launches V4.1 Flash, Using a 552-Billion-Parameter Architecture to Cut Inference Costs
DeepSeek is debuting its new architecture in a low-priced Flash model, leaving real-world task costs to determine whether the efficiency gain holds.
DeepSeek Harness 0.1.5 Arrives With Model Optimization and Expanded Plugin File Workflows
DeepSeek is pushing model upgrades into the agent runtime at the same time, giving developers a broader product interface alongside the new model.
Thu, Sep 10
DeepSeek reportedly engages CITIC Securities for a STAR Market IPO, potentially starting this year with deal size undecided
IPO preparations provide a potential financing route. Formal revenue, cost and capital-demand disclosures would reveal more about the economics of the model business.
DeepSeek plans to route V4 Pro requests to V4.1 Flash after launch and charge Flash rates
Backend substitution reduces migration work but leaves developers needing version identification and regression tests. Lower bills and stable task performance need to be accepted together.
Wed, Sep 9
DeepSeek opens a limited V4.1 Flash beta, claiming a new architecture and native multimodality without a full technical release
The limited beta offers a window to test the new architecture; improvements in capability, speed and cost each need measurement on the same real tasks.
DeepSeek cuts Flash pricing from noon September 10, with cached input at RMB0.02 per million tokens and peak rates still double
Caching receives the largest cut, but actual savings depend jointly on the cache-hit rate, output length and share of work performed at peak times.
DeepSeek seeks about 150 senior backend engineers to rebuild infrastructure under growing user and agent demand
User growth puts backend engineering in the foreground; the payoff from hiring should appear in more reliable task delivery, not merely a larger team.
Mon, Sep 7
DeepSeek to run inference on 160,000 Ascend 950DT chips in Ulanqab while training stays on Nvidia
Training on Nvidia, inference on domestic silicon — where that line sits is the real progress bar for China's compute self-sufficiency.
Sat, Sep 5
DeepSeek said to plan a cluster of more than 160,000 Huawei Ascend 950DT chips in Inner Mongolia
160,000 Ascend chips under one lab's name moves domestic compute from "usable" to "bet on by a leading lab."
Tue, Sep 1
Z.ai's first-half revenue jumps 400% to 954M yuan, still short of its own $200M projection
Revenue up fourfold, stock down sixty percent — that gap is the market repricing the story told by China's model companies.
DeepSeek open-sources V4-Flash-Vision-Exp weights, 305B parameters under an MIT license
305B weights under MIT put the first squeeze not on OpenAI but on every domestic multimodal API billing by the call.
Fri, Aug 28
DeepSeek to raise $7.4 billion at a $74 billion valuation, aimed at compute buildout
Once the racks are self-built, DeepSeek's cost advantage stops being an algorithm question and becomes a capex question.
Thu, Aug 27
Zhipu Open-Sources GLM-5.3-Flash; All Public Beta Traffic Ran on Domestic Chip Clusters
The watershed this time isn't benchmark scores — it's that a frontier-scale model has, for the first time, proven it can survive without NVIDIA's supply rhythm.
Alibaba Open-Sources Qwen3.8-Flash-Next — 125B Model Activates Only 6B Parameters per Token
The one-ninth training-cost figure says more than any benchmark score about which battle the next-generation Qwen aims to win.
Reuters Exclusive: Moonshot AI Seeks 30% Cut From Big Three Cloud Providers for Kimi K3 Hosting
The revenue-share talks were never just about the money — they're about whether Chinese models get a named, legitimate supplier identity on US clouds.
DeepSeek First Seven Months: $70.7M Revenue, API Gross Margin Hits 82.9%
The 82.9% API gross margin proves the inference business itself can make money; the question has always been who pays for the free half.
Mon, Aug 24
NVIDIA Buys Poolside's Technology and Team for $7 Billion, Builds Its Own Open-Source Model to Rival DeepSeek
A chipmaker building the strongest open-source model itself was never about selling the model — it's about becoming the default option that routes everyone's inference onto its own hardware.
DeepSeek Adjusts API Billing — Entire Weekend Now Billed at Off-Peak Rate
Where the pricing sheet loosens is usually where overcapacity first surfaces.
Sat, Aug 22
OpenAI cuts GPT-5.6 Sol input price to $4 per million tokens, down over 20%
A three-month limited-time price cut is not a promotion — it's OpenAI leaving itself an opening to adjust prices again at any time.
DeepSeek Launches Experimental Multimodal Model V4-Flash-Vision-Exp, Reading Up to 600 Images per Request
Comparison numbers produced on a self-run Harness are only worth as much as how quickly third parties can reproduce them.
Tencent's Next-Gen Model Hy4 Appears in Internal Test Interface, Parameter Count to Exceed Hy3's 295B
With self-developed and open-source models sitting side by side in the internal test interface, Tencent has no intention of selecting just one of the two paths.
Fri, Aug 21
DeepSeek Harness Ships rc.8, Adding Image Input and Installable Sub-Agents
Converting images to text and feeding the result to text-only models paves the way with engineering before model capabilities catch up.
Tue, Aug 18
DeepSeek API Adopts Peak/Off-Peak Pricing; V4 Pro Peak Output Rises to 27 Yuan per Million Tokens
When APIs start billing by peak and off-peak windows, compute officially becomes a utility.
Mon, Aug 17
DeepSeek Open-Sources Agent Framework Harness, Hits 95K GitHub Stars in Two Days
The real objective of open-sourcing a runtime is getting others' engineering habits to take root in your abstraction.
Sat, Aug 15
DeepSeek open-sources agent framework Harness: every component is pluggable
Beyond the model, the agent runtime is becoming the next layer to be conquered by open source.
Fri, Aug 14
DeepSeek Releases V4-Pro at One-Seventh the Price of Kimi K3
List price cut to a tenth, cache prices up tenfold — DeepSeek has turned a price war into a price-structure war and shifted the full burden of price comparison onto developers.
DeepSeek open-sources agent framework Harness with an "everything is a plugin" philosophy
With models free and frameworks open-sourced, the truly scarce assets left in the agent race are usage entry points and dominance of the plugin ecosystem.
Thu, Aug 13
xAI Releases Grok 4.6, Intelligence Index Ties GPT-5.6 Sol, Price Steady at $2
Intelligence gap of 1 point, price gap of 8x — the high-priced tier needs a new justification.
DeepSeek Quietly Lists V4-Pro-0813 with 1M Context and 384K Max Output
The four-times-cheaper flagship model carries its own price-increase warning; cost models need to leave a line for it.
Fable 5 Captures 11.4% of Anthropic Revenue in First Month, Token Volume Just 6%
Usage at 6%, revenue at 11.4% — premium models sell task value, not call counts.
Wed, Aug 12
DeepSeek Registers 'DeepSeek Harness Team' Official Account Under Beijing DeepSeek
Between shipping a model and shipping a product sits an organization that can interview every day and open an official account.
Mon, Aug 10
DeepSeek V4-Flash 0731 Third-Party Benchmarks Are In: Just 13B Active Parameters, Agent Score Beats Its Own Pro Edition
Redoing post-training once was enough to overtake the premium version, showing the bar for agent capability is shifting from "stacking parameters" to "refining the recipe" — precisely the stage where the open-source camp is closing the gap fastest.
Unitree Robotics Opens Subscription Tomorrow: Issue Price 150.80 Yuan, Offline Subscription Multiple 2,618x, DeepSeek in Strategic Placement
DeepSeek's stake in Unitree turns the convergence of "brain" and "body" from a forum topic into an equity structure.
Sat, Aug 8
Unitree Technology's strategic placement list includes DeepSeek; Wang Xingxing says it will receive model architecture and intelligent-computing cluster support
The easiest way for a hardware company to get a brain is to make the brain-maker a shareholder.
Fri, Aug 7
DeepSeek Backend Notice Flags Major API Price Hike, Hitting an Inflection Point in the Low-Price Era of Chinese Large Models
The price slasher is raising prices itself — the ledger behind rock-bottom pricing no longer balances.
MiniMax Open-Sources Next-Gen Multimodal Model H3, Tops Hugging Face Trending Chart
Eight domestic chips adapted on the same day says more about what MiniMax has a grip on than a No. 1 spot on the leaderboard.
Unitree Robotics Reveals Strategic-Placement Roster; DeepSeek Allocated About 141 Million Yuan
A model company is paying for robot-body equity — upstream and downstream in embodied AI are crossing into each other's turf.
Thu, Aug 6
Meta debuts first coding agent Muse Code, with output priced at $4.25 per million tokens
The coding-model price war is on — the first shot lands on competitors' margins, and the ammunition is data.
DeepSeek Restarts Second Funding Round, Plans to Raise RMB 50 Billion at Pre-Money Valuation of ~RMB 500 Billion
RMB 1 trillion of intent chases an RMB 50 billion quota — money was never the scarce asset.
Wed, Aug 5
Alibaba Releases 2.4-Trillion-Parameter Qwen3.8-Max, Says Max-Level Open Weights Next Week
What counts is no longer any one release, but the once-a-month release calendar.
DeepSeek's Updated V4-Flash Matches GLM-5.2 in Benchmarks at One-Tenth the Price
When benchmark parity costs a tenth of the price, the middle-tier API has no story left to tell.
Tue, Aug 4
Artificial Analysis Estimates: DeepSeek V4-Flash Costs 3 Cents per Task
Per-million-token pricing is losing its reference value; what counts is the total bill for getting a task done.
Mon, Aug 3
DeepSeek V4-Flash official release goes live, output price cut to 2 yuan per million tokens — one-third of Pro
The moment a cheap model catches up to a premium one, what gets eliminated isn't the premium model — it's model routing as a business.
All Top Five in OpenRouter's Weekly Calls Are Chinese Models; Xiaomi MiMo-V2.5 Takes Top Spot with 10.5 Trillion Tokens
Open source isn't a margin concession — it's moving the distribution channel from someone else's console into your own hands.
Sat, Aug 1
DeepSeek Ships V4-Flash Stable: Weights Open-Sourced Same Day, Agent Benchmarks Overtake Its Own Pro Preview
Run the same model twice, and the second pass is worth 10 extra points — post-training is becoming a more valuable asset than parameter count.
Xiaohongshu Plans $2.2B, 600 MW Data Center in Ulanqab, Inner Mongolia
Power can be bought; cards may not be — two different ledgers.
Zhipu Relaunches GLM Coding Plan Subscription, Moves to Credit-Based Billing Starting at ¥118/Month
One side is handing out weights for free, the other is more than doubling subscription prices — both companies are betting on the same thing: the money isn't in the model.
// related topics — frequently covered alongside #DeepSeek