Codex Reset
AI快讯
GitHub

QQ·微信群

CODEX / SIGNAL STUDIO

你的 Codex,尽在掌握。

#deeplearning-ai

DeepLearning.AI 在 The Batch 中详解 DeepSeek 缓存缩减与 Flash 基准测试表现

DeepLearning.AI 在最新一期 The Batch 中分析了 DeepSeek 的架构,指出其每个 token 的缓存降至 890 字节,较 DeepSeek-V1 缩小了 437 倍。分析还提到,当输入从 4K 扩展至 1M token 时,每个输出 token 的计算量增加了 25%,并指出 Flash 在 AA Index 基准测试中得分超过 V4-Pro。