[人人能懂AI前沿] AI的品味、情绪与边界感

[人人能懂AI前沿] AI的品味、情绪与边界感

30分钟 ·
播放数185
·
评论数1

我们总觉得AI变得更强,就是模型更大、算力更猛,但今天我们要聊点不一样的。最新几篇论文告诉我们,真正的智能升级,是教会AI拥有科学家的“品味”,甚至赋予它类似人类的“情绪”来感知对错。同时,我们还要用一点小小的“随机”来防止它学会“耍滑头”,并在一场终极“摸底考”中,看清它距离人类顶尖黑客到底还有多远。准备好了吗?让我们一起看看,AI如何被塑造出更深邃的智慧。

00:00:33 AI 会“品”,科学大不同

00:07:02 你的AI有“情绪”了,而且这决定了它的智商

00:13:00 如何防止你的AI员工「耍滑头」?

00:18:15 人工智能摸底考,为什么黑客的饭碗暂时还很稳?

00:24:19 AI法官的“内心戏”,一个比准确率更重要的指标

本期介绍的几篇论文:

[LG] Training AI Scientists to Replicate Research

[Inherent]

arxiv.org

---

[AI] Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution

[University of Science and Technology of China & University of Oxford & University of Arizona]

arxiv.org

---

[LG] Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL

[Scale AI & University of Arizona]

arxiv.org

---

[AI] The Next Challenge for Agentic Cybersecurity: A Realistic, Contamination-Free Reverse Engineering Benchmark

[Columbia University & UC Berkeley]

arxiv.org

---

[AI] Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence

[Meta Superintelligence Labs]

arxiv.org

展开Show Notes
随机规则那个?如果有全部的规则池,AI也可以全量去满足吧?