【赞助商】
OpenClaw快报
每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论
传送门 www.xiaoyuzhoufm.com
【目录】
本期的 15 篇论文如下:
[] 🤖 Atria Dawn: The Dawn of Agentic Superintelligence(Atria Dawn:智能体超级智能的黎明)
[] 🎬 Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation(Vidu S2:实时交互、可编辑与空间视频生成)
[] 🤖 ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search(ZGCM-1:一个完全开放且极其高效、面向数学与智能体搜索的基础模型)
[] 🧠 Dream-RSI: Recursive Self-Improvement through Evolving Worlds(Dream-RSI:通过演化世界实现递归自我改进)
[] 🤖 PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models(PhysBrain 1.5:从视觉语言模型到物理基础模型)
[] 🗜 Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction(分组值注意力:通过按需键重建实现高效KV缓存)
[] 🎬 LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows(LynnReal-Omni:面向智能体视觉工作流的原生多模态视频生成)
[] 🤖 RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments(RSIAgent:新环境中面向递归自我改进的自主探索)
[] 🔬 Discovery Foundation Models: Toward Open-Ended Discovery Intelligence(发现基础模型:迈向开放式发现智能)
[] 🎬 BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender(BVB:在 Blender 中通过程序化重建对智能体视频理解进行基准测试)
[] ⚖ How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus(无损投机解码究竟有多无损?数值精度在 Orthrus 中的作用)
[] 🌐 AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video(AlayaVista:从全景状态到透视视频的流式世界建模)
[] 🧩 Kaininja: Extending Native 3D Generators to the Part Level(Kaininja:将原生3D生成器扩展至部件级)
[] 🧠 Omni-Streaming Thinking(全模态流式思维)
[] 🖥 LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents(LLaDA-UI:将块级扩散引入视觉语言 GUI 智能体)

【关注我们】
您还可以在以下平台找到我们,获得播客内容以外更多信息
小红书: AI速递
