SemiAnalysis · @SemiAnalysis_ · X·2026-09-10 19:56·47分钟前
AI 导读

DeepSeek 发布 DeepSeek-V4.1-Flash,采用 552B 骨干的因果编码器-解码器架构,prefill 激活 8B 参数、decode 激活 16B 参数。模型通过稀疏查找访问 196B Engram 记忆,并借助 bounded replay 使持久 KV 缓存比 V4-Flash 少约 8 倍。

SemiAnalysis@SemiAnalysis_
67AI 编辑部评分,满分 100
2026-09-10 19:56· 47分钟前
AI 导读

DeepSeek 发布 DeepSeek-V4.1-Flash,采用 552B 骨干的因果编码器-解码器架构,prefill 激活 8B 参数、decode 激活 16B 参数。模型通过稀疏查找访问 196B Engram 记忆,并借助 bounded replay 使持久 KV 缓存比 V4-Flash 少约 8 倍。

Congrats to @deepseek_ai on releasing DeepSeek-V4.1-Flash!

552B backbone, with a causal encoder-decoder activating just 8B params at prefill, 16B at decode 196B Engram memory accessed through sparse lookups ~8x less persistent KV than V4-Flash through bounded replay