@latentspacepod: In this episode, @OpenAI Chief Research Officer @markchen90 joins @allenpark to flambé shrimp, cook Korean stew, and ch…
Summary
Latent Space podcast hosts OpenAI Chief Research Officer Mark Chen to discuss scaling laws, pre-training, the evals crisis, and OpenAI's research roadmap while cooking.
View Cached Full Text
Cached at: 06/27/26, 07:54 AM
In this episode, @OpenAI Chief Research Officer @markchen90 joins @allenpark to flambé shrimp, cook Korean stew, and chat about being at the frontier of AI research: why scaling laws and pre-training still matter, how OpenAI chooses research bets and allocates compute, what it means to develop research taste, why evals are in crisis, how to avoid benchmark-maxing, and what it will take for models to handle long-horizon real-world work, multimodal reasoning, and eventually end-to-end AI research.
Timestamps: 0:00 Intro 0:28 The Soup Story 1:52 From Trading to AI Research 3:21 How to Develop Research Taste 5:23 RL, Evals, and Superhuman Benchmarks 8:17 Cooking Begins on the Impulse Stove 8:53 Scaling Laws, Pre-Training, and Reasoning 12:33 OpenAI’s Research Roadmap and Compute Allocation 15:48 What Makes a Great Researcher 19:33 The Evals Crisis and Benchmark-Maxing 24:34 Jagged Intelligence, Context, and Long-Horizon Learning 27:14 Shrimp Flambé and New Research Bets 31:32 Multimodal Models and One Architecture 32:36 Vibe Researching and End-to-End AI Research 34:36 Failed Bets, Postmortems, and OpenAI’s Alpha 37:07 Final Taste Test 37:53 Overrated vs. Underrated AI Research 41:00 Closing
Similar Articles
@arcane_bloom: https://x.com/arcane_bloom/status/2077733957995733062
Discusses a Latent Space podcast episode where Anjney Midha explains why AI labs with unlimited GPUs still fail, drawing on his experience at amppublic and a16z.
@OpenAI: Listen to the OpenAI Podcast on— Spotify https://open.spotify.com/show/0zojMEDizKMh3aTxnGLENP… Apple https://podcasts.a…
OpenAI announces the availability of their podcast on major streaming platforms including Spotify, Apple Podcasts, and YouTube.
@natolambert: New podcast with @finbarrtimbers! We survey the latest post-training recipes, from GLM 5.1, Kimi K2.6, DeepSeek V4, Xia…
Nathan Lambert and Finbarr Timbers discuss the latest post-training recipes for large language models, including DeepSeek V4, GLM 5.1, Kimi K2.6, and the industry shift to multi-teacher on-policy distillation.
@EthanHe_42: In @latentspacepod podcast, I shared my view on video generation, world models, LLMs, agents, continual learning and wh…
Ethan He shares his insights from a Latent Space podcast, discussing key ideas about video generation, world models, LLMs, agents, continual learning, and the next frontiers in AI.
@OpenAI: In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of rese…
Terence Tao and Mark Chen discuss how AI is changing mathematical research, from literature search to code generation, and the need to adapt workflows.