Humans&: Bridging IQ and EQ in Machine Learning with Eric Zelikman
Episode
36 min
Read time
2 min
Topics
Productivity, Startups, Fundraising & VC
AI-Generated Summary
Key Takeaways
- ✓STaR Algorithm Scaling: The Self-Taught Reasoner trains models by having them generate solutions iteratively, learning only from correct answers while progressively solving harder problems. N-digit multiplication experiments showed no obvious plateau as training iterations increased, suggesting genuine scalability in reasoning capabilities.
- ✓Model Intelligence Gaps: Current models excel at closed-form verifiable problems like physics or math when given proper context, but fail at understanding long-term implications of their responses. They treat each conversation turn as independent, never asking clarifying questions or expressing uncertainty about user goals.
- ✓Task-Centric Training Limitations: Benchmarks focus on single-task performance for credit assignment between teams rather than measuring how models affect people's lives over time. This paradigm prevents models from learning memory, proactive behavior, or understanding how individual requests fit into broader user contexts and objectives.
- ✓Human-AI Collaboration Advantage: Models that understand individual goals and coordinate with large groups will likely solve fundamental problems faster than autonomous AI working alone for extended periods. Empowering people to pursue their passions grows economic potential rather than simply replacing existing GDP segments with automation.
What It Covers
Eric Zelikman discusses his AI research on reasoning and reinforcement learning at Stanford and XAI, then explains his new company Humansand's mission to build models that understand human goals and collaborate effectively rather than replace people.
Key Questions Answered
- •STaR Algorithm Scaling: The Self-Taught Reasoner trains models by having them generate solutions iteratively, learning only from correct answers while progressively solving harder problems. N-digit multiplication experiments showed no obvious plateau as training iterations increased, suggesting genuine scalability in reasoning capabilities.
- •Model Intelligence Gaps: Current models excel at closed-form verifiable problems like physics or math when given proper context, but fail at understanding long-term implications of their responses. They treat each conversation turn as independent, never asking clarifying questions or expressing uncertainty about user goals.
- •Task-Centric Training Limitations: Benchmarks focus on single-task performance for credit assignment between teams rather than measuring how models affect people's lives over time. This paradigm prevents models from learning memory, proactive behavior, or understanding how individual requests fit into broader user contexts and objectives.
- •Human-AI Collaboration Advantage: Models that understand individual goals and coordinate with large groups will likely solve fundamental problems faster than autonomous AI working alone for extended periods. Empowering people to pursue their passions grows economic potential rather than simply replacing existing GDP segments with automation.
Notable Moment
Zelikman reveals that Google researchers explained task-centric benchmarks persist partly because they enable resource allocation between teams based on percentage improvements, not because they measure what actually matters for helping users accomplish meaningful goals over time.
You just read a 3-minute summary of a 33-minute episode.
Get No Priors: Artificial Intelligence | Technology | Startups summarized like this every Monday — plus up to 2 more podcasts, free.
Pick Your Podcasts — FreeKeep Reading
More from No Priors: Artificial Intelligence | Technology | Startups
Building an Autonomous Delivery Experience with DoorDash Co-Founders Andy Fang and Stanley Tang
Jul 23 · 49 min
Eye on AI
#335 Sriram Raghavan: Why IBM Is Betting Everything on Small AI Models
Apr 19
More from No Priors: Artificial Intelligence | Technology | Startups
Travel Through the Lens of AI with with Booking.com CEO Glenn Fogel
Jul 9 · 41 min
Eye on AI
#324 Sharon Zhou: Inside AMD's Plan to Build Self-Improving AI
Feb 27
More from No Priors: Artificial Intelligence | Technology | Startups
We summarize every new episode. Want them in your inbox?
Building an Autonomous Delivery Experience with DoorDash Co-Founders Andy Fang and Stanley Tang
Travel Through the Lens of AI with with Booking.com CEO Glenn Fogel
How Nuclear Will Unlock Energy Abundance with Valar Atomics Founder Isaiah Taylor
Why Traditional Benchmarks Fail Modern AI Models with OpenAI Research Scientist Noam Brown
Re-engineering the Semiconductor Supply Chain with Intel CEO Lip Bu Tan
Similar Episodes
Related episodes from other podcasts
Eye on AI
Apr 19
#335 Sriram Raghavan: Why IBM Is Betting Everything on Small AI Models
Eye on AI
Feb 27
#324 Sharon Zhou: Inside AMD's Plan to Build Self-Improving AI
The TWIML AI Podcast
Jan 29
The Evolution of Reasoning in Small Language Models with Yejin Choi - #761
20VC (20 Minute VC)
Dec 1
20VC: Scale, Surge, Turing, Mercor: Who Wins & Who Loses in Data Labelling | Is Revenue in Data Labelling Real or GMV? | Why 99% of Knowledge Work Will Go and What Happens Then? | Why SaaS is Dead in a World of AI with Jonathan Siddharth @ Turing
Latent Space
Jul 16
🔬 The Lab of the Future Should Feel Like a Data Center — Andy Beam & Rafa Gómez-Bombarelli, Lila Sciences
Explore Related Topics
This podcast is featured in Best AI Podcasts (2026) — ranked and reviewed with AI summaries.
Read this week's Startups & Product Podcast Insights — cross-podcast analysis updated weekly.
You're clearly into No Priors: Artificial Intelligence | Technology | Startups.
Every Monday, we deliver AI summaries of the latest episodes from No Priors: Artificial Intelligence | Technology | Startups and 192+ other podcasts. Free for one show.
Start My Monday DigestNo credit card · Unsubscribe anytime