← 返回快讯

快讯

推文称 DeepSeek 将智能体记忆缓存压缩至每 token 890 bytes

DeepLearningAI

推文称,针对智能体“读多写少”的场景,DeepSeek 缩小了其记忆缓存:每 token 缓存为 890 bytes,较 DeepSeek-V1 小 437 倍;同时称从 4K 到 1M 输入,每个输出 token 的计算量增加 25%。推文还称 Flash 在 AA Index 上以 39 比 36 领先 V4-Pro,任务成本为 0.27 美元对 0.67 美元。以上均为原文主张,未独立核实。

所属事件 →

事件来源

查看原文
Agents read more than they write, so DeepSeek shrank what they remember. Read the technical breakdown in this week's issue of The Batch. 📰🗜️ 📉 890 bytes of cache per token, 437x smaller than DeepSeek-V1 ⚡ +25% compute per output token from 4K to 1M input 🏆 Flash beats V4-Pro: 39 vs 36 on AA Index, $0.27 vs $0.67/task https://hubs.la/Q04ztt000 #DeepLearningAI #LLMs #AIAgents

前后事件