DeepSeek V4.1 Flash 资讯汇编DeepSeek V4.1 Flash: Everything Confirmed, Reported, and Rumored

资讯汇编12 分钟阅读更新于 2026 年 9 月 10 日

官方通知原文(掘金逐字贴出,最可靠的一段):

◆知微•资讯汇编 · 12 分钟阅读 · 2026 年 9 月 10 日
News Roundup12 min readUpdated September 10, 2026

Original official notice (posted verbatim by Juejin, the most reliable passage):

◆知微•News Roundup · 12 min read · September 10, 2026
DeepSeek V4.1 Flash 资讯汇编

采集时间:2026-09-09 夜间(模型尚未正式发布)

状态:内测版 deepseek-v4.1-flash-expires-on-0910 运行中,9 月 10 日到期;正式版计划 9 月 10 日前后上线

可信度三档:【官方确认】(官网/公开文档可查)|【平台通知转引】(登录态通知,媒体逐字转引)|【社区/未证实】

原始档案见 调研-v41flash-官方信源.md、调研-v41flash-社区与媒体.md

零、三句话说清这次是什么事

  1. 不是单纯降价,是新模型发布。V4.1 Flash 计划 9 月 10 日前后上线,官方称其在性能、费用、速度、总用时上全面超越 V4 Pro。
  2. 最狠的是自动路由。V4.1 Flash 上线后、V4.1 Pro 发布前,指向 V4 Pro 的 API 请求会被全部切到 V4.1 Flash,并按 Flash 的更低单价计费——用户不用改一行代码,模型变强、钱变少。
  3. 9 月 10 日 12:00 的调价是为新模型让路。缓存命中输入降 60%(0.05→0.02 元),未命中降 33%,输出降 11%。

二、关键事实速查

项结论可信度备注
------------
正式模型名deepseek-v4.1-flash(无 -beta)平台通知转引内测名 = 正式名 + -expires-on-0910
上线时间"北京时间 2026 年 9 月 10 日前后"平台通知转引非确切时点,不可写死
架构新的模型结构,原生多模态平台通知转引区别于 vision-exp 的外挂式视觉
上下文 / 输出 / 并发 / 思考模式官方未说明未查到现有 v4-flash 为 1M / 384K / 2500 并发【官方确认】;V4.1 内测限流 20 并发
内测形态base_url 不变,改模型名即用;计费同 V4-Flash;单账号 20 并发;9/10 到期平台通知转引正式版 V4-Flash 并发是 2500
基准分数DeepSeek 未发布任何 V4.1 Flash 基准表官方确认(查无)与 Vision-Exp 发布即附基准的做法相反

官方通知原文(掘金逐字贴出,最可靠的一段):

"DeepSeek 计划于北京时间 2026 年 9 月 10 日前后正式发布 V4.1 Flash 模型。经内部、外部多方测试,V4.1 Flash 在性能、费用、速度、总用时等各项指标上已全面超越 V4 Pro。秉持着对用户负责的态度,在 V4.1 Flash 正式上线之后、V4.1 Pro 上线之前,我们会将对 V4 Pro 的请求全部路由到 V4.1 Flash,并按 V4.1 Flash 单价计费。如您在 V4 Pro 和 V4.1 Flash 的对比测试中发现任何问题,请及时向我们反馈,感谢您的支持!"

三、定价对照(元 / 百万 tokens)

档位现行空闲价【官方确认】新空闲价(9/10 12:00 起)降幅新高峰价
---------------
输入(缓存命中)0.050.02-60%0.04
输入(缓存未命中)1.51.0-33%2.0
输出4.54.0-11%8.0

峰谷规则不变【官方确认】:高峰 = 周一至周五 9:00–12:00、14:00–18:00,高峰价 = 空闲价 ×2。 作为对照,V4-Pro 现行价:缓存命中 / 未命中 / 输出 = 0.15 / 4.5 / 13.5(空闲),高峰 0.30 / 9.0 / 27.0。

注意口径:降 60% 的只是缓存命中档。这也是昨天那篇号外的标题口径,需在文中点明。

四、V4 Pro 自动路由(本次最值得写的点)

  • 政策:V4.1 Flash 上线后、V4.1 Pro 发布前,指向 V4 Pro 的 API 请求全部路由至 V4.1 Flash,按 Flash 单价计费
  • 用户侧:无需任何代码改动(服务端路由)
  • 边界:仅在过渡期;V4.1 Pro 发布后政策如何,官方未说明
  • 官方理由:"秉持着对用户负责的态度"
  • 风险提示:该政策未出现在任何公开文档页面上,只在登录态通知里

按现行价算,同一请求从 V4 Pro 切到 V4.1 Flash(假设新价与 Flash 一致),输出档从 13.5 元降到 4 元,未命中输入从 4.5 元降到 1 元——这是"升级 + 降价"同时发生的原因。

五、社区实测:一致的方向 vs 相反的声音

一致的方向:确实更快

速度数字散布极大,均非官方口径,引用时必须带上"谁测的、什么场景":

数字来源场景
---------
284 vs 97 tok/s上海证券报转述社区V4.1 vs 现 V4,媒体转述
355(峰值 365)、TTFT 178msX @riba2534(KOCPC 转引)对比 V4 Pro:吞吐 5.7 倍、首字延迟降 77%
507、420、328智东西汇总 X 晒图不同任务,最高值
300–350(峰值近 500)掘金转 Reddit常规文本流式
260 / 200–300 / 400+小红书、B站一天内散布 200→600

稳妥写法:只说"社区普遍反馈比 V4-Flash 明显更快",具体数字必须标注来源与场景,且要说明差异来自峰谷时段、任务长度、测试方法。

相反的声音(务必收录,别只挑好听的)

  1. "居然比上代贵"(个人 14 组任务、约 3 亿 token 实测):同款三关小游戏,V4.1 端到端 34.5 分钟 vs V4 30.5 分钟——吐字快 ≠ 交活快,省下的生成时间又被工具调用验证吃掉(等工具 18.7 分钟 vs 6.6 分钟)
  2. 第一财经(9/8):开发者反馈"5 分钟 10 块钱""瞬间几十块没了",对"成本更低"持保留;推测更快 = 单位时间吞更多 token
  3. 过度思考 / 翻车(r/LocalLLaMA):负载重时频繁失败、文件编辑失败、4M tokens 仍未完成一个 Frogger 游戏;七奇迹基准耗时 2 小时、花费 2.60 美元(竞品约 30 分钟)
  4. 自认知混乱:有人问"你是谁",它回答自己是 Claude——被视为"没当正式版本看"的证据
  5. 能力不均:综合网页任务("浏览器操作系统")拉胯,桌面资源加载异常;火箭模拟物理表现仍落后顶级竞品
  6. 内测性质:B站测评指出暂不支持视频、只测了图片;部分第三方适配没跟上,甚至直接显示不支持图片

正面案例(可用)

  • 3D"S 形路线停车":V4.1 约 25.4 秒零碰撞完成;V4 用 11 分 27 秒且主画面上下颠倒
  • 小店任务(Node.js + 购物车 + 优惠码 + 幂等下单):V4.1 用 1 分 41 秒,V4 用 5 分 54 秒
  • 单社区编程对比:V4.1 总时长 18 分 49 秒 / 11.59M tokens,vs Vision-Exp 30 分 11 秒 / 20.31M tokens(快 38%、省 43% token)——但注明 V4.1 有任务未跑完
  • 多模态无幻觉:博主"向阳乔木"发西装照,模型答"条纹西装",放大确认确实如此

六、媒体在怎么解读

共识:几乎一致解读为"用 Flash 的价格和速度去替代 Pro 的能力"的降维打击,叠加同步降价,属罕见的"升级降价同步"策略。

各家角度:

媒体角度
------
上海证券报定性为行业罕见的"升级降价同步";内测问卷"能否全面替代 V4 Pro"暴露野心
钱江晚报·潮新闻聚焦"里子"——架构换代、原生多模态、迁移成本近零;限流 20 vs 2500 说明是功能验证版
第一财经最早质疑"成本更低";补全 V4 系列时间线
智东西社区速度数字整合,"快得飞起"
虎嗅放在"低成本训练 ≠ 低资金需求"框架下,关联 IPO 与 500 亿私募
KOCPC"Flash 价格、Pro 野心"的商业化解读
科创板日报调价 + 路由 + IPO 三线并进

分歧点:集中在"成本是否真降"——第一财经与个人实测持保留态度。

七、时间线(已核对)

日期事件
------
2026-04-24DeepSeek V4 Preview 发布
2026-07-31V4-Flash-0731 正式版 API 上线公测,Agent 能力增强
2026-08-13V4-Pro-0813 正式版 + 大幅调价;同日 DeepSeek Harness 开发者预览版开源
2026-08-21V4-Flash-Vision-Exp 视觉模型上线(外挂式)
2026-08-26网易有道 LobsterAI 上线 DSH
2026-09-07DeepSeek Harness 团队公开招聘 150 个岗位
2026-09-08V4.1 Flash 内测开启(限流 20、9/10 到期),发放"能否替代 Pro"问卷
2026-09-09上海证券报 / IT之家 / 第一财经 / 智东西等多家报道;路透披露 IPO 筹备
2026-09-10(计划)V4.1 Flash 正式发布;12:00 起 Flash 新价生效;V4 Pro 请求路由切换;内测接口下线

注意:这次降价距离 8 月 13 日那次大幅涨价还不到一个月。

八、DSH 生态反应(本号关注)

  • 网易有道 LobsterAI 9 月 8 日宣布率先集成 V4.1 Flash 内测版,是生态里动作最快的一个
  • 未查到 DSH 官方针对 deepseek-v4.1-flash 的专属适配,也未见"dsh-web 梁神模式 flash"的专门动作。现有 DSH / 梁神模式资料均围绕 V4 Flash 与 V4 Pro
  • DSH 默认通道机制(V4-native 协议下 deepseek-v4-flash 作默认高吞吐通道、难题自动升级 deepseek-v4-pro)意味着:一旦 V4.1 Flash 上线,默认通道与升级路径都可能随之改变,这是 DSH 用户真正该关注的点
  • 官方 awesome-deepseek-agent 已覆盖 20 款工具(Claude Code、Cline、Codex、OpenCode、Roo 等)的路由配置
  • 未查到 Claude Code / Cline / Roo 社区针对 V4.1 Flash 的专门讨论帖

九、IPO 与资本动态(全部"媒体报道,未经官方证实")

路透社 9/9 援引两名知情人士:DeepSeek 已聘请中信证券筹备科创板 IPO,希望年内启动流程;上市时间、募资规模、发行估值均未定;DeepSeek 与中信证券均未回应置评。

其他媒体口径(均无官方确认):

  • 当前估值约 5000 亿元;上市后市值区间或落在 1.5 万亿–2.5 万亿元
  • 4 月首次开放融资;6 月首轮交割募资超 500 亿元、投后超 3500 亿元(梁文锋个人约 200 亿、腾讯 100 亿、宁德时代体系 50 亿、京东/网易/IDG 各 30 亿)
  • 7 月中旬启动二轮,投前估值升至 5000 亿元,两轮累计募资超 1000 亿元
  • 2026 年前七个月营收约 4.75 亿元、净亏 7.15 亿元(The Information / 彭博)

写作建议:这条与模型本身无关,且完全未经官方证实。若在推文中提及,必须写"据路透社报道""未经官方证实",且不要与模型发布建立因果联系。

十、写作取材建议

可以放心写的

  • 9 月 10 日 12:00 起的 Flash 新价目、降幅、峰谷规则(有官方通知 + 现行牌价可对照)
  • V4 Pro 自动路由政策(有通知原文,注明"官方通知称")
  • 官方"全面超越 V4 Pro"的表述(标注为官方口径,不转述成自己的评价)
  • 内测形态:限流 20、9/10 到期、改模型名即用

要小心处理的

  • 上线时间写"9 月 10 日前后",不要写"9 月 10 日 12:00 发布"(12:00 是调价时间,不是发布时间)
  • 速度数字:必须"社区实测 + 谁测的 + 什么场景",且说明数字散布 200–600
  • "成本更低":官方口径与部分开发者体感相反,建议两面都给
  • 规格(上下文/输出/并发):官方没说,别填

不能写的

  • 把"梁文锋回归/复出"当事实("梁圣回来了"只能是社区玩梗的转述)
  • 把 IPO 写成既成事实
  • 把社区跑分当官方基准
  • 把内测版(20 并发、9/10 下线)与正式版混为一谈

一个提醒:昨天那篇号外只写了调价,漏掉了新模型发布与自动路由这个更大的新闻。9 月 10 日 12:00 前有必要发一篇续篇,或者至少在原文评论区/后续推送里补上——否则读者会以为这只是一次普通降价。

附:主要信源清单

官方:api-docs.deepseek.com(价格页 / 更新日志 / news260813)、platform.deepseek.com(登录态通知) 媒体转引:掘金(逐字通知全文)、上海证券报、IT之家、第一财经、钱江晚报·潮新闻、智东西、虎嗅、巨亨网、KOCPC、科创板日报、路透社 社区:什么值得买、微博长文、APPSO、知乎、小红书、B站、X(@riba2534、@MiaAI_lab)、Reddit r/LocalLLaMA(AGI Hunt 转引)、稀土掘金、CSDN、8news.ai、Geeky Gadgets、FrontierNews、GPT Proto

Collected: night of 2026-09-09 (the model has not been officially released yet)

Status: beta build deepseek-v4.1-flash-expires-on-0910 running, expiring September 10; release version planned for around September 10

Three credibility tiers: 【officially confirmed】 (verifiable on the official site / public docs) | 【platform-notice relay】 (logged-in notice, relayed verbatim by media) | 【community/unverified】

Original archives: 调研-v41flash-官方信源.md, 调研-v41flash-社区与媒体.md

0. The Whole Story in Three Sentences

配图
  1. This isn't just a price cut — it's a new model launch. V4.1 Flash is slated to go live around September 10, and the company says it comprehensively surpasses V4 Pro in performance, cost, speed, and total elapsed time.
  2. The wildest part is the auto-routing. After V4.1 Flash goes live and before V4.1 Pro launches, API requests aimed at V4 Pro will all be switched to V4.1 Flash and billed at Flash's lower rates — users don't change a line of code, the model gets stronger and the money gets smaller.
  3. The 12:00 repricing on September 10 clears the way for the new model. Cache-hit input drops 60% (¥0.05→¥0.02), misses drop 33%, output drops 11%.

2. Key Facts at a Glance

ItemConclusionCredibilityNote
------------
Official model namedeepseek-v4.1-flash (no -beta)Platform-notice relayBeta name = official name + -expires-on-0910
Launch timing"around September 10, 2026, Beijing time"Platform-notice relayNot an exact moment; don't write it as fixed
ArchitectureA new model structure, natively multimodalPlatform-notice relayDiffers from vision-exp's bolt-on vision
Context / output / concurrency / thinking modeNot stated officiallyNot foundCurrent v4-flash is 1M / 384K / 2500 concurrency【officially confirmed】; V4.1 beta caps at 20 concurrent
Beta formbase_url unchanged, just change the model name; billed the same as V4-Flash; 20 concurrent per account; expires 9/10Platform-notice relayOfficial V4-Flash concurrency is 2500
Benchmark scoresDeepSeek has published no V4.1 Flash benchmark tableOfficially confirmed (nothing found)The opposite of Vision-Exp, which shipped with benchmarks

Original official notice (posted verbatim by Juejin, the most reliable passage):

"DeepSeek plans to officially release the V4.1 Flash model around September 10, 2026, Beijing time. After extensive internal and external testing, V4.1 Flash has comprehensively surpassed V4 Pro across every metric, including performance, cost, speed, and total elapsed time. In the spirit of an attitude of responsibility toward users, after V4.1 Flash officially goes live and before V4.1 Pro launches, we will route all requests to V4 Pro to V4.1 Flash and bill them at V4.1 Flash rates. If you find any problems in comparative testing between V4 Pro and V4.1 Flash, please report them to us promptly. Thank you for your support!"

3. Pricing Comparison (yuan / million tokens)

TierCurrent off-peak price【officially confirmed】New off-peak price (from 9/10 12:00)ReductionNew peak price
---------------
Input (cache hit)0.050.02-60%0.04
Input (cache miss)1.51.0-33%2.0
Output4.54.0-11%8.0

Peak/off-peak rules unchanged【officially confirmed】: peak = Monday–Friday 9:00–12:00 and 14:00–18:00; peak price = off-peak price ×2. For comparison, V4-Pro's current prices: cache hit / miss / output = 0.15 / 4.5 / 13.5 (off-peak), peak 0.30 / 9.0 / 27.0.

Mind the framing: only the cache-hit tier drops 60%. That was also the headline framing of yesterday's news flash, and it needs to be spelled out in the piece.

4. V4 Pro Auto-Routing (the Most Newsworthy Point This Time)

  • Policy: after V4.1 Flash goes live and before V4.1 Pro launches, all API requests aimed at V4 Pro are routed to V4.1 Flash and billed at Flash rates
  • User side: no code changes required (server-side routing)
  • Boundaries: interim period only; what happens after V4.1 Pro launches is unstated
  • Official reason: "an attitude of responsibility toward users"
  • Risk note: the policy appears on no public documentation page, only in a logged-in notice

At current prices, for the same request switching from V4 Pro to V4.1 Flash (assuming the new prices match Flash's), the output tier drops from ¥13.5 to ¥4 and missed input from ¥4.5 to ¥1 — that's why "upgrade + price cut" happen at the same time.

5. Community Tests: One Consistent Direction vs. Dissenting Voices

The Consistent Direction: It Really Is Faster

The speed numbers are scattered wildly and none are official; when citing them you must include "who tested it and in what scenario":

NumberSourceScenario
---------
284 vs 97 tok/sShanghai Securities News relaying the communityV4.1 vs current V4, media relay
355 (peak 365), TTFT 178msX @riba2534 (relayed by KOCPC)vs V4 Pro: 5.7× throughput, 77% lower time to first token
507, 420, 328Zhidx roundup of X screenshotsDifferent tasks, top values
300–350 (peak near 500)Juejin relaying RedditOrdinary streaming text
260 / 200–300 / 400+Xiaohongshu, BilibiliSpread from 200→600 within a single day

The safe way to write it: say only that "the community broadly reports it's noticeably faster than V4-Flash"; any specific number must carry its source and scenario, plus a note that the differences come from peak/off-peak hours, task length, and test method.

Dissenting Voices (Must Include — Don't Cherry-Pick Only the Good News)

  1. "It's actually more expensive than the previous generation" (one individual's 14 task tests, ~300 million tokens): for the same three-level mini game, V4.1 took 34.5 minutes end to end vs 30.5 minutes for V4 — fast tokens ≠ fast delivery; the generation time saved was eaten by tool-call verification (18.7 minutes waiting on tools vs 6.6 minutes)
  2. Yicai (9/8): developers reported "¥10 in 5 minutes" and "dozens of yuan gone in an instant," and it withheld judgment on "lower cost," speculating that faster = more tokens swallowed per unit of time
  3. Overthinking / breakdowns (r/LocalLLaMA): frequent failures under heavy load, failed file edits, 4M tokens and still no finished Frogger game; the Seven Wonders benchmark took 2 hours and cost $2.60 (competitors take about 30 minutes)
  4. Self-identity confusion: asked "who are you," it said it was Claude — taken as evidence it "isn't being treated as a release version"
  5. Uneven abilities: general web tasks ("browser operating system") fall flat, desktop resource loading misbehaves; rocket-simulation physics still trails top competitors
  6. Beta nature: a Bilibili review notes video isn't supported yet and only images were tested; some third-party integrations haven't caught up and even report no image support at all

Positive Cases (Usable)

  • 3D "S-shaped path parking": V4.1 finished in about 25.4 seconds with zero collisions; V4 took 11 minutes 27 seconds and rendered the main view upside down
  • Small shop task (Node.js + shopping cart + coupon codes + idempotent checkout): V4.1 took 1 minute 41 seconds, V4 took 5 minutes 54 seconds
  • One community coding comparison: V4.1 total 18 minutes 49 seconds / 11.59M tokens, vs Vision-Exp 30 minutes 11 seconds / 20.31M tokens (38% faster, 43% fewer tokens) — though note that V4.1 left some tasks unfinished
  • No multimodal hallucination: the blogger "Xiangyang Qiaomu" (向阳乔木) posted a suit photo and the model answered "pinstripe suit"; zooming in confirmed it was right

6. How the Media Are Reading It

Consensus: almost uniformly read as a dimensional-strike play — "use Flash's price and speed to replace Pro's capability" — compounded by a simultaneous price cut, a rare "upgrade and price cut at once" strategy.

How each outlet frames it:

OutletAngle
------
Shanghai Securities NewsCalls it an industry-rare "upgrade and price cut at once"; the beta survey question "can it fully replace V4 Pro" exposes the ambition
Qianjiang Evening News · Chao NewsFocuses on the substance — architecture generational change, native multimodality, near-zero migration cost; the 20 vs 2500 concurrency cap shows it's a feature-validation build
YicaiFirst to question "lower cost"; fills in the V4 series timeline
ZhidxAggregates community speed numbers, "fast as lightning"
HuxiuFrames it under "low-cost training ≠ low capital needs," tying in the IPO and the ¥50 billion private placement
KOCPCA commercialization read: "Flash price, Pro ambition"
STAR Market DailyRepricing + routing + IPO advancing on three fronts

Point of divergence: centered on "whether costs really drop" — Yicai and individual testers withhold judgment.

7. Timeline (Verified)

DateEvent
------
2026-04-24DeepSeek V4 Preview released
2026-07-31V4-Flash-0731 release-version API opens for public beta, Agent capabilities enhanced
2026-08-13V4-Pro-0813 release version + major repricing; DeepSeek Harness developer preview open-sourced the same day
2026-08-21V4-Flash-Vision-Exp vision model goes live (bolt-on)
2026-08-26NetEase Youdao LobsterAI launches on DSH
2026-09-07DeepSeek Harness team opens 150 job postings to public hiring
2026-09-08V4.1 Flash beta opens (20 concurrency cap, expires 9/10); "can it replace Pro" survey distributed
2026-09-09Shanghai Securities News / ITHome / Yicai / Zhidx and others report; Reuters discloses IPO preparations
2026-09-10 (planned)V4.1 Flash official release; new Flash prices effective from 12:00; V4 Pro request routing switches over; beta endpoint taken offline

Note: this price cut comes less than a month after the major price hike on August 13.

8. DSH Ecosystem Reaction (What We're Watching)

  • NetEase Youdao LobsterAI announced on September 8 that it was the first to integrate the V4.1 Flash beta, the fastest mover in the ecosystem
  • We found no DSH official support specifically for deepseek-v4.1-flash, and no dedicated "dsh-web Liang God mode flash" move either. Existing DSH / Liang God mode materials all center on V4 Flash and V4 Pro
  • DSH's default channel mechanism (under the V4-native protocol, deepseek-v4-flash as the default high-throughput channel, auto-escalating to deepseek-v4-pro on hard problems) means: once V4.1 Flash ships, both the default channel and the escalation path may change with it — this is the point DSH users should really care about
  • The official awesome-deepseek-agent already covers routing configuration for 20 tools (Claude Code, Cline, Codex, OpenCode, Roo, and others)
  • We found no dedicated discussion threads about V4.1 Flash in the Claude Code / Cline / Roo communities

9. IPO and Capital Moves (All "Media Reports, Not Officially Confirmed")

Reuters on 9/9, citing two people familiar with the matter: DeepSeek has engaged CITIC Securities to prepare a STAR Market IPO and hopes to start the process this year; listing timing, fundraising size, and offering valuation are all undecided; neither DeepSeek nor CITIC Securities responded to requests for comment.

Other outlets' accounts (none officially confirmed):

  • Current valuation around ¥500 billion; post-listing market cap could land in the ¥1.5–2.5 trillion range
  • Opened financing for the first time in April; the first round closed in June, raising over ¥50 billion at a post-money valuation above ¥350 billion (Liang Wenfeng personally about ¥20 billion, Tencent ¥10 billion, the CATL ecosystem ¥5 billion, JD.com / NetEase / IDG ¥3 billion each)
  • Second round launched in mid-July, pre-money valuation rising to ¥500 billion, cumulative fundraising across both rounds above ¥100 billion
  • Revenue of about ¥475 million and a net loss of ¥715 million in the first seven months of 2026 (The Information / Bloomberg)

Writing advice: this item has nothing to do with the model itself and is entirely unconfirmed officially. If you mention it in a post, you must write "according to Reuters" and "not officially confirmed," and do not draw a causal link to the model release.

10. Sourcing Advice for Writing

Safe to write

  • The new Flash price sheet effective 12:00 on September 10, the reductions, and the peak/off-peak rules (backed by an official notice + current published rates for comparison)
  • The V4 Pro auto-routing policy (notice text available; note "the official notice says")
  • The official "comprehensively surpasses V4 Pro" phrasing (label it as the official line; don't restate it as your own assessment)
  • Beta details: 20 concurrency cap, expires 9/10, change the model name and it works

Handle with care

  • Write the launch timing as "around September 10," not "released at 12:00 on September 10" (12:00 is the repricing time, not the release time)
  • Speed numbers: must be "community test + who tested + what scenario," and note that the numbers range from 200 to 600
  • "Lower cost": the official line conflicts with some developers' experience — give both sides
  • Specs (context / output / concurrency): the company hasn't said, so don't fill them in

Must not write

  • Treating "Liang Wenfeng's return/comeback" as fact ("Saint Liang is back" can only be relayed as a community joke)
  • Writing the IPO as a done deal
  • Treating community benchmark runs as official benchmarks
  • Conflating the beta build (20 concurrency, offline 9/10) with the release version

One reminder: yesterday's news flash covered only the price cut and missed the bigger news of the new model launch and auto-routing. Before 12:00 on September 10, it's worth publishing a follow-up, or at least adding it in the comments on the original piece or a later push — otherwise readers will think this was just an ordinary price cut.

Appendix: Main Source List

Official: api-docs.deepseek.com (pricing page / changelog / news260813), platform.deepseek.com (logged-in notices) Media relays: Juejin (verbatim notice text), Shanghai Securities News, ITHome, Yicai, Qianjiang Evening News · Chao News, Zhidx, Huxiu, CNYES, KOCPC, STAR Market Daily, Reuters Community: SMZDM, Weibo long-form posts, APPSO, Zhihu, Xiaohongshu, Bilibili, X (@riba2534, @MiaAI_lab), Reddit r/LocalLLaMA (relayed by AGI Hunt), Rare Earth Juejin, CSDN, 8news.ai, Geeky Gadgets, FrontierNews, GPT Proto