@Dinosaur_liu: 给全部中美模型厂商说个鬼故事 DeepSeek V4 Flash 只有284b参数 激活参数只有区区13b

X AI KOLs Timeline 模型

摘要

据称 DeepSeek V4 Flash 仅有2840亿参数,激活参数仅130亿,对中美模型厂商构成效率冲击。

给全部中美模型厂商说个鬼故事 DeepSeek V4 Flash 只有284b参数 激活参数只有区区13b https://t.co/xGWUEQH1Qc
查看原文
查看缓存全文

缓存时间: 2026/08/03 09:38

给全部中美模型厂商说个鬼故事

DeepSeek V4 Flash 只有284b参数

激活参数只有区区13b https://t.co/xGWUEQH1Qc

相似文章

@Fenng: 看到自媒体写的这么一段儿「最新的第四代 WeLM-80B,总参数已经只有 800 亿了,激活 30 亿,激活率只有 3.75%。作为对比——国内极致成本性能的代表 DeepSeek-V4-Flash,总参数 2840 亿、激活 130 亿…

X AI KOLs Timeline

Fenng shares a self-media comparison between the fourth-generation WeLM-80B (80B total params, 3B activated, 3.75% activation rate) and DeepSeek-V4-Flash (284B total, 13B activated, 4.6% activation rate), with a humorous comment.

@FuckAnthropic: 做了对比分析,综合看,DeepSeek V4 Flash-0731 大致是一个 Opus 4.7 - 4.8 之间水平的模型,以极小激活规模进入了前沿 Agent 模型竞争区间,用约 1/12~1/60 的 token 成本进入前沿 Ag…

X AI KOLs Timeline

作者对比分析认为,DeepSeek V4 Flash-0731 在极小激活规模下达到 Opus 4.7-4.8 水平,以极低 token 成本进入前沿 Agent 模型区间,整体超越 GLM-5.2,但短板仍是困难仓库级编程和长时程工程。

DeepSeek-V4-Flash-0731

Product Hunt

DeepSeek 宣布推出 DeepSeek-V4-Flash-0731,这是一款以 Flash 级别定价提供先进能力的前沿智能体模型。