2026.09.15 | 智能体超级智能黎明;实时可编辑空间视频生成

2026.09.15 | 智能体超级智能黎明;实时可编辑空间视频生成

12分钟 ·
播放数93
·
评论数0

【赞助商】
OpenClaw快报
每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论
传送门 www.xiaoyuzhoufm.com

【目录】
本期的 15 篇论文如下:

[00:30] 🤖 Atria Dawn: The Dawn of Agentic Superintelligence(Atria Dawn:智能体超级智能的黎明)
[01:13] 🎬 Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation(Vidu S2:实时交互、可编辑与空间视频生成)
[01:58] 🤖 ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search(ZGCM-1:一个完全开放且极其高效、面向数学与智能体搜索的基础模型)
[02:47] 🧠 Dream-RSI: Recursive Self-Improvement through Evolving Worlds(Dream-RSI:通过演化世界实现递归自我改进)
[03:24] 🤖 PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models(PhysBrain 1.5:从视觉语言模型到物理基础模型)
[04:06] 🗜 Grouped Value Attention: Efficient KV Caching via On-Demand Key Reconstruction(分组值注意力:通过按需键重建实现高效KV缓存)
[04:51] 🎬 LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows(LynnReal-Omni:面向智能体视觉工作流的原生多模态视频生成)
[05:45] 🤖 RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments(RSIAgent:新环境中面向递归自我改进的自主探索)
[06:30] 🔬 Discovery Foundation Models: Toward Open-Ended Discovery Intelligence(发现基础模型:迈向开放式发现智能)
[07:13] 🎬 BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender(BVB:在 Blender 中通过程序化重建对智能体视频理解进行基准测试)
[07:56] ⚖ How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus(无损投机解码究竟有多无损?数值精度在 Orthrus 中的作用)
[08:39] 🌐 AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video(AlayaVista:从全景状态到透视视频的流式世界建模)
[09:26] 🧩 Kaininja: Extending Native 3D Generators to the Part Level(Kaininja:将原生3D生成器扩展至部件级)
[10:13] 🧠 Omni-Streaming Thinking(全模态流式思维)
[11:00] 🖥 LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents(LLaDA-UI:将块级扩散引入视觉语言 GUI 智能体)

【关注我们】
您还可以在以下平台找到我们,获得播客内容以外更多信息
小红书: AI速递