2026.09.18 | V4.1-Flash压缩KV提效;SoL-Pi优化智能体降本
HuggingFace 每日AI论文速递

2026.09.18 | V4.1-Flash压缩KV提效;SoL-Pi优化智能体降本

12分钟 120 1周前
节目简介
来源:小宇宙
【赞助商】
OpenClaw快报
每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论
传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43
【目录】
本期的 15 篇论文如下:
[00:32] 🗜 DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression(DeepSeek-V4.1-Flash:将 KV 缓存压缩推向极限)
[01:15] ⚡ SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness(SoL-Pi:递归扩展自动化研究循环以实现高效智能体执行框架)
[01:55] 🛑 When EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation(当 EOS token 不一致:理解在线策略蒸馏中的长度膨胀)
[02:33] 🧪 An Empirical Study of Harness Design for Coding Agents(面向编码智能体的 Harness 设计实证研究)
[03:20] 🌍 JEPA-Anything: Learning Predictive Models across Different Worlds(JEPA-Anything:跨不同世界学习预测模型)
[04:09] 🕵 RiskChainBench: A Benchmark for Obfuscated Platform Message Restoration and Evidence-Grounded Web Investigation(RiskChainBench:面向混淆平台消息还原与证据支撑网络调查的基准)
[04:56] 🎓 RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning(RetireOPD:面向智能体强化学习的自退场在线策略蒸馏)
[05:40] 📄 WeVisDoc: From Coverage to Capability for Robust End-to-End Document Parsing(WeVisDoc:从覆盖到能力,实现鲁棒的端到端文档解析)
[06:29] 🔄 Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents(反思、修订、复用:面向GUI智能体的免训练技能演化)
[07:10] 🤖 VABench: Measuring Embodied Spatial Intelligence through Visual Demonstrations, Active Perception, and Metric Control(VABench:通过视觉演示、主动感知和度量控制测量具身空间智能)
[07:56] 🎥 Video DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation(Video DeltaNet:面向直播视频生成的视频原生混合注意力)
[08:42] 🦾 FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations(FAMOS:基于稀疏观测的前馈式三维铰接建模)
[09:24] 🖼 UFO: Chain-of-Evaluation for Omni-Condition Alignment in Multi-Modal Image Generation(UFO:面向多模态图像生成全条件对齐的评估链)
[10:07] 🌍 Can MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model(MiniMax-H3 能否推理物理世界?一项全模态生成模型评估)
[10:54] 🧠 When2Think: Learning Difficulty-Aware Length Control for Efficient Hybrid Reasoning Models(When2Think:面向高效混合推理模型的难度感知长度控制学习)
【关注我们】
您还可以在以下平台找到我们,获得播客内容以外更多信息
小红书: AI速递

加入我们的 Discord

与播客爱好者一起交流

立即加入

扫描微信二维码

添加微信好友,获取更多播客资讯

微信二维码

播放列表

自动播放下一个

播放列表还是空的

去找些喜欢的节目添加进来吧