“We need a field of Reward Function Design” by Steven Byrnes

08/12/2025 10 min

Listen "“We need a field of Reward Function Design” by Steven Byrnes"

Descargar episodio Ver en sitio original

Episode Synopsis

(Brief pitch for a general audience, based on a 5-minute talk I gave.) Let's talk about Reinforcement Learning (RL) agents as a possible path to Artificial General Intelligence (AGI) My research focuses on “RL agents”, broadly construed. These were big in the 2010s—they made the news for learning to play Atari games, and Go, at superhuman level. Then LLMs came along in the 2020s, and everyone kinda forgot that RL agents existed. But I’m part of a small group of researchers who still thinks that the field will pivot back to RL agents, one of these days. (Others in this category include Yann LeCun and Rich Sutton & David Silver.) Why do I think that? Well, LLMs are very impressive, but we don’t have AGI (artificial general intelligence) yet—not as I use the term. Humans can found and run companies, LLMs can’t. If you want a human to drive a car, you take an off-the-shelf human brain, the same human brain that was designed 100,000 years before cars existed, and give it minimal instructions and a week to mess around, and now they’re driving the car. If you want an AI to drive a car, it's … not that. [...] ---Outline:(00:15) Let's talk about Reinforcement Learning (RL) agents as a possible path to Artificial General Intelligence (AGI)(02:17) Reward functions in RL(04:23) Reward functions in neuroscience(05:25) We need a (far more robust) field of reward function design(06:06) Oh man, are we dropping this ball(07:30) Reward Function Design: Neuroscience research directions(08:14) Reward Function Design: AI research directions(08:46) Bigger picture ---
First published:
December 8th, 2025

Source:
https://www.lesswrong.com/posts/oxvnREntu82tffkYW/we-need-a-field-of-reward-function-design
---
Narrated by TYPE III AUDIO.
---Images from the article:Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

More episodes of the podcast LessWrong (30+ Karma)

“Announcing RoastMyPost” by ozziegooen 17/12/2025

“The Bleeding Mind” by Adele Lopez 17/12/2025

“Towards training-time mitigations for alignment faking in RL” by Vlad Mikulik, Hoagy, Joe Benton, Benjamin Wright, Jonathan Uesato, Monte M, Fabien Roger, evhub 17/12/2025

“Still Too Soon” by Gordon Seidoh Worley 17/12/2025

“Non-Scheming Saints (Whether Human Or Digital) Might Be Shirking Their Governance Duties, And, If True, It Is Probably An Objective Tragedy” by JenniferRM 17/12/2025

“Mistakes in the Moonshot Alignment Program and What we’ll improve for next time” by Kabir Kumar 17/12/2025

“Dancing in a World of Horseradish” by lsusr 17/12/2025

[Linkpost] “Announcing: MIRI Technical Governance Team Research Fellowship” by yams, peterbarnett, Aaron_Scher, Robi Rahman 17/12/2025

“Radiology Automation Does Not Generalize to Other Jobs” by Xodarap 16/12/2025

“GPT-5.2 Is Frontier Only For The Frontier” by Zvi 16/12/2025

Ver todos los episodios

ZARZA We are Zarza, the prestigious firm behind major projects in information technology.

“We need a field of Reward Function Design” by Steven Byrnes

Listen "“We need a field of Reward Function Design” by Steven Byrnes"

Episode Synopsis

More episodes of the podcast LessWrong (30+ Karma)

Deep web or Invisible Internet

Email on your own domain, luxury or need?

Bandwidth: Broadband or Narrowband?

Personnel recruitment via Web

Deep web or Invisible Internet

Subdomains, a glance with the experts!

Free Internet, a prediction in Nostradamus style

Educational Technology: From traditional to digital

Localhost, there’s no place like 127.0.0.1

Googling with breathtaking tricks you ignore

Gray Hat Hacking, those with ambiguous ethics…

Internet Predators on the prowl

Dot COM: The Internet’s dominant TLD