searchlore

Back to John Schulman
John Schulman (OpenAI Cofounder) — Reasoning, RLHF, & plan for 2027 AGI
John Schulman

John Schulman (OpenAI Cofounder) — Reasoning, RLHF, & plan for 2027 AGI

May 15, 2024

Key Takeaways


John Schulman on how posttraining tames the shoggoth, and the nature of the progress to come...

𝐄𝐏𝐈𝐒𝐎𝐃𝐄 𝐋𝐈𝐍𝐊𝐒
* Apple Podcasts: https://podcasts.apple.com/us/podcast/john-schulman-openai-cofounder-reasoning-rlhf-plan/id1516093381?i=1000655679622
* Spotify: https://open.spotify.com/episode/1ivzHH9RWciXe4O1rKtldf?si=53503781e05f4d8f
* Transcript: https://www.dwarkeshpatel.com/p/john-schulman/
* Me on Twitter: https://twitter.com/dwarkesh_sp/

𝐒𝐏𝐎𝐍𝐒𝐎𝐑
* CommandBar is an AI user assistant that any software product can embed to non-annoyingly assist, support, and unleash their users. Used by forward-thinking CX, product, growth, and marketing teams. Learn more at https://www.commandbar.com/

If you’re interested in advertising on the podcast, fill out this form: https://airtable.com/appxGOvFLDLP5dlzv/pagFVrbHRohW6F2bZ/form

𝐓𝐈𝐌𝐄𝐒𝐓𝐀𝐌𝐏𝐒
00:00:00 - Pre-training, post-training, and future capabilities
00:17:20 - Plan for AGI 2025
00:29:43 - Teaching models to reason
00:40:10 - The Road to ChatGPT
00:51:33 - What makes for a good RL researcher?
01:00:18 - Keeping humans in the loop
01:14:36 - State of research, plateaus, and moats

Video