X Post: https://x.com/himanshustwts/status/1963218210531999846
Kalomaze: https://x.com/kalomaze/status/1963202667234091437
TIMESTAMPS:
00:00:00 - TEASER
00:01:06 - INTRO
00:01:49 - Organic Influence
00:02:46 - "Extremely Silly Jester"
00:04:04 - School Days, Diving into ML Research
00:09:27 - Developing "Research Taste" & The Art of Selection
00:11:45 - Dropping College, Parents' Reaction, Plan Ahead
00:18:13 - How a Reddit Post Led to a Job at Shopify
00:23:53 - Why He Chose Prime Intellect Over Other Offers
00:29:27 - A Day in the Life: Workload at PI
00:31:53 - Story of Verifiers, GRPO, and Semi-Verifiable Rewards
00:40:02 - Progress on Synthesizers
00:43:28 - The Environment Hub: A Hugging Face for RL?
00:47:47 - How ideas to research emerge, Experimentation
00:51:38 - Defining "Taste" in Research
00:54:15 - The Future of Supervised Fine-Tuning
01:01:48 - The Flaw in "Pure Reasoner" Models & Thoughts on GPT-OSS
01:05:34 - Perspective on the Chinese Open Source AI Surge
01:07:58 - The Art of Training a Good RL Model
01:10:08 - How to Learn a New Field: A Hacker's Approach
01:12:33 - Handling Disagreements with Formally Experienced Researchers 01:14:49 - Recipe to RLHF
01:19:47 - Scaling Bigger Models vs. Designing Better Rewards
01:26:31 - Rethinking Progress & The Post-AGI Narrative
01:29:27 - The Most Unexpected Part of Working at Prime Intellect
01:30:09 - Trivia Round Begins
01:30:25 - Code Reviews from Will?
01:33:43 - A Hidden Secret About Will!
01:34:26 - The Underrated Mindset That Gives Him an Edge
01:36:18 - Take on AI Tech Twitter ("TPOT")
01:38:16 - How Kalomaze shaped up + Advise
01:41:17 - Podcast Wrap-up

Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.