Album

HuggingFace 每日AI论文速递

10分钟速读热门AI论文

拨号上网 佚名
1.45万 订阅 644 集 3天前
播客简介
每天10分钟,带您快速了解当日HuggingFace热门AI论文内容。每个工作日更新,欢迎订阅。 📢播客节目在小宇宙、Apple Podcast平台搜索【HuggingFace 每日AI论文速递】 🖼另外还有图文版,可在小红书搜索并关注【AI速递】
节目
2026.07.07 | 跨平台智能体学习新范式;科研构思可复用技能提炼

2026.07.07 | 跨平台智能体学习新范式;科研构思可复用技能提炼

HuggingFace 每日AI论文速递

【赞助商】 OpenClaw快报 每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论 传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43 【目录】 本期的 14 篇论文如下: [00:32] 🤖 UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning(UI-MOPD:面向持续GUI智能体学习的多平台在线策略蒸馏) [01:30] 💡 ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes(ResearchStudio-Idea:基于证据的科研构思技能套件——来自机器学习会议成果) [02:21] 🎨 PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space(PixWorld:在像素空间中统一3D场景生成与重建) [03:11] 🧩 OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers(OmniOpt:现代优化器的分类、几何结构与基准测试) [04:04] 🤖 GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation(GigaWorld-1:构建用于机器人策略评估的世界模型路线图) [04:55] 🧩 Vision Pretraining for Dense Spatial Perception(面向密集空间感知的视觉预训练) [05:54] 🤖 EVA-Client: A Unified Data Collection, Inference, and Deployment Framework for Embodied Policies on Real Robots(EVA-Client:面向实体机器人上的具身策略的统一数据收集、推理与部署框架) [06:47] 🎥 Wan-Streamer v0.2: Higher Resolution, Same Latency(Wan-Streamer v0.2:更高分辨率,相同延迟) [07:48] 🤖 InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization(InternVLA-A1.5:统一理解、潜在预知与动作以实现组合泛化) [08:46] 🔍 Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-Language Retrieval(所有视觉标记都同等重要吗?面向视觉-语言检索的保留对象证据的标记合并方法) [09:59] 🧠 KVpop -- Key-Value Cache Compression with Predictive Online Pruning(KVpop——基于预测性在线剪枝的键值缓存压缩) [10:50] 🧠 dOPSD: On-Policy Self-Distillation for Diffusion Language Models(dOPSD:扩散语言模型的在线自蒸馏方法) [11:41] 🎨 Perceptual Flow Matching for Few-Step Generative Modeling(感知流匹配:用于少步生成建模) [12:37] 🔬 Multi-Turn Agentic Scientific Literature Search via Workflow Induction(多轮交互式科学文献搜索:基于工作流归纳的方法) 【关注我们】 您还可以在以下平台找到我们,获得播客内容以外更多信息 小红书: AI速递

14分钟
61
3天前
2026.07.06 | 策略对齐破解推理盲区;轻量外挂实现实时修正

2026.07.06 | 策略对齐破解推理盲区;轻量外挂实现实时修正

HuggingFace 每日AI论文速递

【赞助商】 OpenClaw快报 每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论 传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43 【目录】 本期的 9 篇论文如下: [00:35] 🎯 The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning(优化训练策略的幻影:单调推理策略作为大语言模型强化学习的真正目标) [01:34] 🛠 VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon(VLA-Corrector:一种用于自适应动作视界的轻量级检测与校正推理框架) [02:35] 🤖 Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots(Embodied.cpp:面向异构机器人的具身AI模型便携推理运行时) [03:43] 🔢 OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers(OrbitQuant:面向图像和视频扩散Transformer的数据无关量化方法) [04:38] 📊 DataComp-VLM: Improved Open Datasets for Vision-Language Models(DataComp-VLM:改进的视觉语言模型开放数据集) [05:36] 🛡 Securing the AI Agent: A Unified Framework for Multi-Layer Agent Red Teaming(保护AI智能体:一种多层次智能体红队测试的统一框架) [06:36] ☁ Interpretation-Oriented Cloud Removal via Observation-Anchored Residual Flow with Geo-Contextual Alignment(面向解译的云去除:基于观测锚定残差流与地理上下文对齐的方法) [07:38] 📄 MultAttnAttrib: Training-Free Multimodal Attribution in Long Document Question Answering(多注意力归因:长文档问答中的无训练多模态归因) [08:25] 🔗 AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation(AGE:面向图检索增强生成的自适应掩码图嵌入方法) 【关注我们】 您还可以在以下平台找到我们,获得播客内容以外更多信息 小红书: AI速递

9分钟
59
4天前
2026.07.03 | 小模型本地化击败大模型;自主策略演化聚焦结构合成

2026.07.03 | 小模型本地化击败大模型;自主策略演化聚焦结构合成

HuggingFace 每日AI论文速递

【赞助商】 OpenClaw快报 每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论 传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43 【目录】 本期的 15 篇论文如下: [00:31] 🧩 Program-as-Weights: A Programming Paradigm for Fuzzy Functions(程序即权重:面向模糊函数的编程范式) [01:24] 🧠 EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments(EvoPolicyGym:在交互环境中评估自主策略演化) [02:24] 🧠 AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents(AgenticSTS:面向长时程LLM智能体的有界内存测试平台) [03:18] 🔍 Morphing into Hybrid Attention Models(变形为混合注意力模型) [04:07] 📊 AgenticDataBench: A Comprehensive Benchmark for Data Agents(AgenticDataBench:面向数据智能体的综合性基准测试) [05:12] ⚡ Multi-Resolution Flow Matching: Training-Free Diffusion Acceleration via Staged Sampling(多分辨率流匹配:通过分阶段采样的无训练扩散加速) [05:52] 🎬 WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory(世界导演:构建具有持久动态记忆的可控世界模拟器) [06:49] 🏥 Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning(打破失败级联:面向医学多模态推理的步骤感知强化学习) [07:37] 🎨 Optimizing Visual Generative Models via Distribution-wise Rewards(通过分布级奖励优化视觉生成模型) [08:32] 🎯 SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use(SkillCoach:用于评估和增强智能体技能使用的自我演化评分标准) [09:21] 🖐 AGVBench: A Reliability-Oriented Benchmark of Data Augmentation for Vein Recognition(AGVBench:面向静脉识别的可靠性导向数据增强基准) [10:16] 🔬 From SRA to Self-Flow: Data Augmentation or Self-Supervision?(从SRA到Self-Flow:数据增强还是自监督?) [11:10] 🧠 Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads(对数几率贡献评分识别非字面检索头) [11:58] 🎯 AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models(AnyGroundBench:面向视觉语言模型中视频定位的专业领域基准) [12:46] 🤔 When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search(搜索代理何时应提问:面向澄清感知的深度搜索基准DiscoBench) 【关注我们】 您还可以在以下平台找到我们,获得播客内容以外更多信息 小红书: AI速递

14分钟
99+
1周前
评价

空空如也

加入我们的 Discord

与播客爱好者一起交流

立即加入

扫描微信二维码

添加微信好友,获取更多播客资讯

微信二维码

播放列表

自动播放下一个

播放列表还是空的

去找些喜欢的节目添加进来吧