Create

Sign in to ReadmeX

or
ReadmeX
ReadmeX

A clearer picture, in a conversation.

Catch up on what matters, then ask a little deeper.

Your community and people briefings stay personal to you.

Story

ByteDance Seed paper ties DeepSeek long-context swings to KV-cache phase sensitivity

AI summary

Per DoNews, ByteDance's Seed team posted an arXiv paper in late September 2026 arguing that "phase sensitivity" introduced by chunked KV-cache compression is the main driver of long-context retrieval fluctuations in DeepSeek-V4 series models. Covering open-weight models including DeepSeek-V4-Flash, V4-Pro and V4.1-Flash, the work says compression windows impose token phase coordinates, so the same information can be retrieved with accuracy differing by as much as 40 percentage points across phases — a periodic degradation that averaged benchmark scores tend to hide. The team calls for better cache-compression designs to improve long-context stability.

Why it matters: If it holds, long-context evaluation and deployment need to track phase-dependent stability, not just average benchmark scores.

ByteDance SeedDeepSeekDeepSeek-V4-Flash

9
Source textDoNews · 4 min read

首页快讯商业Image 3 消费游戏3C家电汽车文娱Image 4 体育专栏直播专题

APP Image 5

Image 7: user

Image 8: user

首页快讯商业Image 9 消费游戏3C家电汽车文娱Image 10 体育专栏直播专题APP

DoNews>快讯>字节Seed团队揭示DeepSeek模型相位敏感性问题

字节Seed团队揭示DeepSeek模型相位敏感性问题

2026-10-09 09:03:03

5621

分享到

2026年9月底,字节跳动Seed团队在arXiv发布论文,指出分块KV缓存压缩技术引发的‘相位敏感性’是DeepSeek-V4系列模型长上下文检索性能波动的主因。研究覆盖DeepSeek-V4-Flash、V4-Pro及V4.1-Flash等开源模型,发现因压缩窗口引入Token相位坐标,导致相同信息在不同相位下检索准确率差异高达40个百分点。该现象造成周期性性能衰减,易被平均基准分数掩盖。团队呼吁优化缓存压缩设计以提升长上下文稳定性。

免责声明:本文内容由开放的智能模型自动生成,仅供参考。

最新文章

Image 16: 谷歌云发布 Gemini Agent,定位“通用工作智能体” 商业 谷歌云发布 Gemini Agent,定位“通用工作智能体” 谷歌云发布Gemini智能体,定位通用工作AI,支持问答、知识处理、内容生成、编程等,可作个人助理或团队成员,兼容多模型并提供API,现处私有预览。 杨亮 11小时前

Image 17: 尊界回应制动踏板支架底座相关问题:提供免费升级选择 商业 尊界回应制动踏板支架底座相关问题:提供免费升级选择 尊界汽车称制动踏板支架底座无实际断裂故障,但将免费升级该部件并复核设计验证与质量管理,以提升安全冗余。 杨亮 11小时前

Image 18: Biohub牵头打造AI制药数据基建,规模增至18亿美元 商业 Nous Research完成B轮融资,估值15亿美元 Nous Research完成9000万美元B轮融资,估值15亿美元;开源智能体Hermes克隆超2400万次,占全球AI token用量2.5%。 杨亮 13小时前

Image 19: Biohub牵头打造AI制药数据基建,规模增至18亿美元 商业 AI算力初创公司Lambda拟在IPO前最后一轮融资中筹集40亿美元 Lambda拟IPO前融资最多40亿美元,估值145亿美元;黑石与Coatue领投;目标2027年上市;积压订单从6月150亿增至9月500亿美元。 杨亮 13小时前

Image 20: Biohub牵头打造AI制药数据基建,规模增至18亿美元 商业 Mistral AI 发布 Mistral Large 4 模型公开预览版 Mistral AI发布万亿参数多模态模型ML4(le Chonk),支持160+语言,欧盟合规训练,性能达开源SOTA,部分领域超闭源模型。 杨亮 13小时前

Image 21: Biohub牵头打造AI制药数据基建,规模增至18亿美元 商业 OpenAI一次性公开377项AI数学成果 OpenAI发布377项AI生成数学成果,涵盖多领域;响应AGMAI建议,承诺规范发布流程、提升质量与透明度。 杨亮 13小时前

Image 22: Biohub牵头打造AI制药数据基建,规模增至18亿美元 商业 Biohub牵头打造AI制药数据基建,规模增至18亿美元 Biohub扩大18亿美元“虚拟生物学倡议”,联合美能源部、NIH、DeepMind、Isomorphic Labs、Meta及NVIDIA。 杨亮 14小时前

Image 23: Liquid AI 发布 Open d1 系列:3B 模型实现零输出 token 商业 Liquid AI 发布 Open d1 系列:3B 模型实现零输出 token Liquid AI发布零输出token实时决策模型Open d1:d1-3B(多模态)与d1-omni-600M(文本+图像/音频)。 杨亮 14小时前

Image 24

关于我们| 电子协议| 合作联系| 京ICP备2025120072号

网站信息Image 29

Copyright © DoNews 2000-2026 All Rights Reserved
京ICP备2025120072号
联系地址:北京市海淀区宝盛东路兴华绿色产业楼3层307室(东升地区)
邮箱:[email protected]
网上有害信息举报专区: www.12377.cn

Copyright © DoNews 2000-2026 All Rights Reserved
京ICP备2025120072号

Image 30京公网安备11010802023059号

Image 31

Image 32

获取验证码 获取验证码

登录

密码登录 >

验证码登录 >

登录即代表您已阅读同意用户协议和隐私协议

#热门搜索

热门文章

最新文章

推荐文章

Read the original →

How we got here

  1. ByteDance Seed ties long-context swings to chunked KV cache 'phase sensitivity'AIbase AI新闻 · DeepSeek
  2. China's tech giants raise ~$56B in debt and equity to pre-buy AI capacity36氪 人工智能 · DeepSeek
  3. Study: Only 3.6% of 857 Chinese frontier AI releases disclosed safety resultsSemiAnalysis · DeepSeek
  4. UK ICO secures data protection changes from 10 AI developers, targets agentsThe Next Web · DeepSeek
  5. Zhipu's GLM release cycle is turning Chinese LLM makers into 'product-cycle stocks'36氪 人工智能 · DeepSeek
  6. Chinese gaming giants pour billions into AI, with NetEase and Tencent backing DeepSeek钛媒体 · DeepSeek

Comments

I've used this: share my experience What I think: share my view
How important is this story?No ratings yet

No comments yet. Start the conversation.