Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Episode
124 min
Read time
3 min
Topics
Investing, Startups, Fundraising & VC
AI-Generated Summary
Key Takeaways
- ✓Claude Code adoption measurement: To verify AI coding adoption claims, scrape GitHub's public commit API for Claude Code's signature sign-off string, then calculate daily counts as a percentage of total GitHub commits. O'Laughlin built a cron job doing exactly this and found Claude Code reached roughly 4% of all GitHub commits within approximately two weeks of tracking — a growth rate he describes as faster than any trend he has previously observed.
- ✓AI as junior analyst framework: Treat Claude Code outputs the way a senior analyst treats a junior analyst's work — useful for aggregating and formatting raw information, but requiring expert review before conclusions are trusted. The critical gap is that current models lack meta-level learning: a human junior analyst accumulates pattern recognition and judgment over repeated cycles, building expertise. Claude Code does not yet compound that experience across sessions, making domain expert oversight non-negotiable.
- ✓Context window hygiene for long tasks: Run complex research tasks within a single one-million-token context window rather than compacting aggressively or splitting across sessions. Use sub-agents for discrete subtasks so each sub-agent maintains a clean context, while the primary window stays uncluttered. Separate the task prompt from the evaluation rubric into distinct steps to reduce sycophantic drift, particularly with Opus 4.6, which tends toward agreement when task and rubric are combined.
- ✓Agent swarms vs. agent teams distinction: Claude's experimental multi-agent team feature underperforms because it lacks reinforcement learning to coordinate context-aware task allocation. By contrast, Gemini 2.5 Flash swarms meaningfully improve output quality. For practical use, sub-agents with clearly scoped tasks and their own context windows outperform agent teams. O'Laughlin used swarms to run internal model benchmarks — running 20 iterations of the same problem set — a workflow previously inaccessible without engineering resources.
- ✓Excel and Bloomberg replacement trajectory: Claude Code using Python and matplotlib already produces higher-quality charts faster than Excel, and the workflow is cheaper. O'Laughlin's firm is actively replacing Bloomberg Terminal data pulls with direct API feeds processed through Claude Code. The underlying argument: Excel and Bloomberg are human-formatted IDEs for information work; once an AI agent can retrieve, analyze, and visualize data directly, the legacy interface layer becomes friction rather than value.
What It Covers
SemiAnalysis co-founder Doug O'Laughlin joins Latent Space to detail how Claude Code transformed his firm's research workflow, tracking AI-generated GitHub commits to quantify adoption, analyzing historical semiconductor memory cycles, and arguing that Claude Code 4.5's ability to one-shot complex multi-step tasks marks a genuine capability threshold for white-collar knowledge work automation.
Key Questions Answered
- •Claude Code adoption measurement: To verify AI coding adoption claims, scrape GitHub's public commit API for Claude Code's signature sign-off string, then calculate daily counts as a percentage of total GitHub commits. O'Laughlin built a cron job doing exactly this and found Claude Code reached roughly 4% of all GitHub commits within approximately two weeks of tracking — a growth rate he describes as faster than any trend he has previously observed.
- •AI as junior analyst framework: Treat Claude Code outputs the way a senior analyst treats a junior analyst's work — useful for aggregating and formatting raw information, but requiring expert review before conclusions are trusted. The critical gap is that current models lack meta-level learning: a human junior analyst accumulates pattern recognition and judgment over repeated cycles, building expertise. Claude Code does not yet compound that experience across sessions, making domain expert oversight non-negotiable.
- •Context window hygiene for long tasks: Run complex research tasks within a single one-million-token context window rather than compacting aggressively or splitting across sessions. Use sub-agents for discrete subtasks so each sub-agent maintains a clean context, while the primary window stays uncluttered. Separate the task prompt from the evaluation rubric into distinct steps to reduce sycophantic drift, particularly with Opus 4.6, which tends toward agreement when task and rubric are combined.
- •Agent swarms vs. agent teams distinction: Claude's experimental multi-agent team feature underperforms because it lacks reinforcement learning to coordinate context-aware task allocation. By contrast, Gemini 2.5 Flash swarms meaningfully improve output quality. For practical use, sub-agents with clearly scoped tasks and their own context windows outperform agent teams. O'Laughlin used swarms to run internal model benchmarks — running 20 iterations of the same problem set — a workflow previously inaccessible without engineering resources.
- •Excel and Bloomberg replacement trajectory: Claude Code using Python and matplotlib already produces higher-quality charts faster than Excel, and the workflow is cheaper. O'Laughlin's firm is actively replacing Bloomberg Terminal data pulls with direct API feeds processed through Claude Code. The underlying argument: Excel and Bloomberg are human-formatted IDEs for information work; once an AI agent can retrieve, analyze, and visualize data directly, the legacy interface layer becomes friction rather than value.
- •Memory cycle regime analysis via AI: O'Laughlin used Claude Code to aggregate NAND and DRAM price data back to the 1980s, combining paid data APIs, SERP search results, and macroeconomic covariates like consumer sentiment and WFE data. He then attempted fine-tuning a Chronos 2 time-series foundation model for price prediction but concluded regime changes — where historical correlations invert — make memory price forecasting via ML unreliable. The residual value was a comprehensive structured dataset and regime-by-regime narrative dashboard built in days rather than months.
- •AI CapEx parallels to railroad build-out: Historical railroad construction consumed 4.8% of GNP annually and represented 25% of total gross fixed capital investment for a sustained decade. Current AI infrastructure spending, including Stargate alone at roughly 2% of US GDP, is on a trajectory to match or exceed that. Railroads produced three distinct boom-bust cycles over 45 years before stabilizing. O'Laughlin expects AI infrastructure to follow multiple cycles rather than one continuous expansion, with demand and supply curves crossing at an unknown future point.
Notable Moment
O'Laughlin revealed that his firm's hiring case study — a multi-step company analysis task used to evaluate research candidates — has been his ongoing benchmark for AI agents since early 2024. His baseline threshold was not human expert performance but simply beating the worst human submissions. Claude Code 4.5 cleared that bar decisively, which he treats as the practical definition of a capability threshold worth acting on.
Episode Transcript
This crap makes mistakes all the time. Mhmm. All the time. It is still just like a Like, I think of it once again as like a junior analyst. Right? The analyst goes and does all this like really pain in the ass information, and you bring it all together to make a good decision at the top. Historically, what happens is that junior analyst, who I once was, went and gathered all that information. And after doing this enough times, there's a meta level thinking that's happening where it's like, okay. Here is what I really understand and how this type of analysis I'm an expert in, actually. I'm very good at. I consistently have a hit rate. Now I'm the expert. Right? I don't think that meta level learning is there yet. We'll see if l ones do it right. Everyone who's spending 1 quadrillion dollars in the world thinks it will. It better it better happen by if you're spending, you know, a trillion dollars and there's not meta level learning. But for me in our firm, that massively amplifies everyone who is an expert. Because, like, you have to still do something that you can't just, like, slop it up. It's very obvious to me what it's slop. Welcome to Lien Space. Yeah. Thank you for having me, man. I after all this time, I just is it okay if I just call you Swyx? I feel like it's the that's that's where my brain is. So I've known you for so long. You can call me Neil if you aren't. I'm not there. You know? Yeah. Yeah. I mean, it's been it's been a long time. It's been it's been a long time coming. I think I first met you at, like, New Orleans or, like, one of the one of the neurophyses. Yeah. Yeah. I met you at one of the neurophyses in Perp I think it was Vancouver. Right? Yeah. No. No. I think it was like some burger party. Yeah. Yeah. Yeah. And you were like, hey. Like, who's this tall, dude? I'm like, woah. Okay. Yeah. Yeah. Well, I mean, it's just like I I knew about you, and we've we've, like, been in that, you know, pen pals for a long time. So it was, like, cool meeting in person. I mean yeah. Yeah. I think that was the first time I ever met you in person. So yeah. Amazing. I I didn't go to the New Orleans when I really wish I did. I love New Orleans, honestly. Yay. So There are two New Orleans just in a row, and, yeah, honestly, we should go back there. Yeah. Are you guys going to Melbourne or the Australia one this year? I have I don't even think that far out. Yeah. But on it, but that sounds pretty interesting to me. I think I can't remember which one. There's a there's something in in in Korea this year. Right? Yeah. …
Get the full transcript (28,909 words) + summary by email — free
One-time email with the complete transcript and AI summary of this episode. No account needed.
One email, no spam. We’ll also show you what SignalCast does.
You just read a 3-minute summary of a 121-minute episode.
Get Latent Space summarized like this every Monday — plus up to 2 more podcasts, free.
Pick Your Podcasts — FreeKeep Reading
More from Latent Space
🔬“We have foundation models for language, not for physics” — Anima Anandkumar, Bren Professor of Computing
Aug 26 · 83 min
a16z Podcast
Can Anyone Catch NVIDIA? | The Future of Chips and Infrastructure
Jul 15
More from Latent Space
Simulation: the new Scaling Law — Joon Sung Park, Simile AI
Aug 21 · 69 min
This Week in Startups
The Startup Building the First Hotel on the Moon…
Jun 15
Books, tools, and gear mentioned in this episode
SignalCast may earn commission on purchases via these links.
Tools
“He then attempted fine-tuning a Chronos 2 time-series foundation model for price prediction but concluded regime changes make memory price forecasting via ML unreliable.”
by GitHub
“scrape GitHub's public commit API for Claude Code's signature sign-off string, then calculate daily counts as a percentage of total GitHub commits.”
- Claude CodeRecommended
by Anthropic
“Claude Code transformed his firm's research workflow, tracking AI-generated GitHub commits to quantify adoption... Claude Code 4.5's ability to one-shot complex multi-step tasks marks a genuine capability threshold for white-collar knowledge work automation.”
by Anthropic
“Separate the task prompt from the evaluation rubric into distinct steps to reduce sycophantic drift, particularly with Opus 4.6, which tends toward agreement when task and rubric are combined.”
by Google
“By contrast, Gemini 2.5 Flash swarms meaningfully improve output quality.”
by Bloomberg
“O'Laughlin's firm is actively replacing Bloomberg Terminal data pulls with direct API feeds processed through Claude Code.”
More from Latent Space
We summarize every new episode. Want them in your inbox?
🔬“We have foundation models for language, not for physics” — Anima Anandkumar, Bren Professor of Computing
Simulation: the new Scaling Law — Joon Sung Park, Simile AI
🔬The BioAI Phase Shift - Matthew McPartlon & Neil Patil, Chai Discovery
The Inference Engineering Masterclass — Philip Kiely & Ali Taha, Baseten
Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI
Similar Episodes
Related episodes from other podcasts
a16z Podcast
Jul 15
Can Anyone Catch NVIDIA? | The Future of Chips and Infrastructure
This Week in Startups
Jun 15
The Startup Building the First Hotel on the Moon…
The Tim Ferriss Show
Apr 7
#860: Daredevil Michelle Khare — How to Become a YouTube Superstar, Open Impossible Doors (FBI, Secret Service, etc.), Craft Jedi-Level Cold Emails, and Use Fear-Setting to Change Your Life
The Joe Rogan Experience
Sep 9
#2551 - Daniel Kokotajlo
a16z Podcast
Sep 9
Who Grades the AI Models? | Ben Horowitz & Rayan Krishnan
Explore Related Topics
This podcast is featured in Best AI Podcasts (2026) — ranked and reviewed with AI summaries.
Read this week's Investing & Markets Podcast Insights — cross-podcast analysis updated weekly.
You're clearly into Latent Space.
Every Monday, we deliver AI summaries of the latest episodes from Latent Space and 192+ other podcasts. Free for one show.
Start My Monday DigestNo credit card · Unsubscribe anytime