This Is How to Tell if Writing Was Made by AI
Episode
48 min
Read time
2 min
Topics
Startups, Marketing, Artificial Intelligence
AI-Generated Summary
Key Takeaways
- ✓AI Detection Accuracy: Pangram Labs achieves a false positive rate of 1 in 10,000 and a false negative rate of roughly 1%, far exceeding the ~90% human baseline accuracy. The model scales beyond simple perplexity metrics by using deep learning trained on millions of side-by-side human and AI writing pairs to detect subtle decision patterns.
- ✓How AI Writing Gets Detected: LLMs make thousands of micro-decisions when constructing even 100 words of text, and their output clusters into a narrow region of all possible writing. Pangram trains a model to recognize these decision patterns through contrast learning — pairing a human review with an AI-generated version of the same content to identify imperceptible differences.
- ✓Active Learning Pipeline: After an initial training pass on known human and AI samples, Pangram scans a larger corpus to surface false positives and false negatives, then feeds those edge cases back into retraining. This self-improving loop continuously pushes the model closer to the human-AI boundary where detection is hardest.
- ✓AI Slop Economics on Reddit: Startups sell services to brands promising organic-seeming AI bot mentions on Reddit, where bots post normal-seeming replies and occasionally name-drop products. This gaming matters because LLMs train on Reddit data, meaning seeded brand mentions in Reddit threads increase the likelihood those brands appear in future AI-generated responses.
- ✓Internet Contamination Scale: Roughly 40% of internet pages are now AI-generated, driven largely by SEO content farms switching to AI to produce keyword-targeting articles at near-zero cost. Medium crossed 50% AI-generated new articles roughly 18 months ago, while Reddit sits at around 10% today, up from 7% a year prior.
What It Covers
Max Spiro, founder of Pangram Labs, explains how his AI detection platform achieves a 1-in-10,000 false positive rate by training deep learning models on tens of millions of paired human and AI writing samples, while approximately 40% of the current internet is already AI-generated content.
Key Questions Answered
- •AI Detection Accuracy: Pangram Labs achieves a false positive rate of 1 in 10,000 and a false negative rate of roughly 1%, far exceeding the ~90% human baseline accuracy. The model scales beyond simple perplexity metrics by using deep learning trained on millions of side-by-side human and AI writing pairs to detect subtle decision patterns.
- •How AI Writing Gets Detected: LLMs make thousands of micro-decisions when constructing even 100 words of text, and their output clusters into a narrow region of all possible writing. Pangram trains a model to recognize these decision patterns through contrast learning — pairing a human review with an AI-generated version of the same content to identify imperceptible differences.
- •Active Learning Pipeline: After an initial training pass on known human and AI samples, Pangram scans a larger corpus to surface false positives and false negatives, then feeds those edge cases back into retraining. This self-improving loop continuously pushes the model closer to the human-AI boundary where detection is hardest.
- •AI Slop Economics on Reddit: Startups sell services to brands promising organic-seeming AI bot mentions on Reddit, where bots post normal-seeming replies and occasionally name-drop products. This gaming matters because LLMs train on Reddit data, meaning seeded brand mentions in Reddit threads increase the likelihood those brands appear in future AI-generated responses.
- •Internet Contamination Scale: Roughly 40% of internet pages are now AI-generated, driven largely by SEO content farms switching to AI to produce keyword-targeting articles at near-zero cost. Medium crossed 50% AI-generated new articles roughly 18 months ago, while Reddit sits at around 10% today, up from 7% a year prior.
Notable Moment
When a researcher attempted to evade Pangram by running AI text through multiple translation layers — English to Chinese to formal Chinese to Hebrew and back to English — the model still correctly identified the output as AI-generated, suggesting the underlying decision patterns survive significant linguistic transformation.
Episode Transcript
Introducing Fidelity Trader Plus, the next generation of advanced trading from Fidelity. Customize your tools and charts and access them seamlessly across desktop, web, and mobile for faster trades anywhere you go. Try the all new Fidelity Trader Plus. Learn more about our most powerful trading platform yet at fidelity.com/traderplus. Investing involves risk, including risk of loss. Fidelity Brokerage Services, LLC, member NYSE SIPC. The thing about AI for business, it may not automatically fit the way your business works. At IBM, we've seen this firsthand. But by embedding AI across HR, IT, and procurement processes, we've reduced cost by millions, slashed repetitive tasks, and freed thousands of hours for strategic work. Now we're helping companies get smarter by putting AI where it actually pays off, deep in the work that moves the business. Let's create smarter business. IBM. Find home wherever you roam at Sonesta ES and Simply Suites, where longer stays feel comfortable, flexible, and easy. Stretch out and enjoy spacious accommodations and home like amenities designed to help you settle in and stay productive or relaxed for however long you need. And when you're a Sonesta Travel Pass member, staying at Sonesta ES and Simply Suites means earning points toward free nights, upgrades, and more with every eligible stay. Go to sonesta.com to book your stay and unlock the best rates with Sonesta travel pass. Here today, roam tomorrow. Join now at sonesta.com. Terms and conditions apply. Bloomberg Audio Studios. Podcasts, radio, news. Hello, and welcome to another episode of the Odd Lots podcast. I'm Jill Wiesenthal. And I'm Tracy Alloway. So, Tracy, you know, you ever come across some writing and you can't articulate exactly why, but you're like, I'm pretty sure AI wrote this. Does this happen too much? So full disclosure, I haven't really thought about it that much. Really? Yeah. Because the thing is, I probably should think about it more. But there's a lot of bad writing out there, and I've become sort of inured to it. And I also think that I don't know. Trying to figure out whether or not something was generated by AI nowadays, if you actually dedicate a lot of your own time to doing that, that is a huge mental burden to be attempting. Especially, you and I are in the journalism industry. How many of the pitches do you think that we get from PRs right now are being generated by AI? Imagine if you're reading each one of those and trying to figure it out on a daily basis. You know what I suppose I think about it the most is, someone will respond to a tweet. Yeah. And I'd be like, well, if this is a real person, then maybe this person deserves some engagement and ask a question or I wanna respond. But if it's a bot in the bot, then obviously, I don't. And that's where I'm like, oh, you know what? I wanna figure it out. I would like to …
Get the full transcript (10,780 words) + summary by email — free
One-time email with the complete transcript and AI summary of this episode. No account needed.
One email, no spam. We’ll also show you what SignalCast does.
You just read a 3-minute summary of a 45-minute episode.
Get Odd Lots summarized like this every Monday — plus up to 2 more podcasts, free.
Pick Your Podcasts — FreeKeep Reading
More from Odd Lots
What the OpenAI-Hugging Face Hack Really Tells Us About AI Danger
Aug 17 · 59 min
Eye on AI
#325 Phelim Brady: Why AI's Future Depends on Human Judgement
Mar 9
More from Odd Lots
A Historic El Niño Is Coming That Could Cost the World Trillions
Aug 14 · 55 min
20VC (20 Minute VC)
20VC: 70% of Neolabs Will Die | There Will be a $100BN US Open-Source Model | Data is a Trillion $ Market | Governments Cannot Regulate Models: It is Too Late | The Cyber Attacks to Come Will be Insane with Anastasios Angelopoulos @ Arena
Aug 3
Books, tools, and gear mentioned in this episode
SignalCast may earn commission on purchases via these links.
Tools
- Pangram LabsBy guest
“Max Spiro, founder of Pangram Labs, explains how his AI detection platform achieves a 1-in-10,000 false positive rate by training deep learning models on tens of millions of paired human and AI writing samples”
More from Odd Lots
We summarize every new episode. Want them in your inbox?
What the OpenAI-Hugging Face Hack Really Tells Us About AI Danger
A Historic El Niño Is Coming That Could Cost the World Trillions
Trucking Is Booming Again, And Drivers Aren't Happy About It
NYT CEO Meredith Kopit Levien on Running a Media Brand in the Age of AI
Introducing: Our Town
Similar Episodes
Related episodes from other podcasts
Eye on AI
Mar 9
#325 Phelim Brady: Why AI's Future Depends on Human Judgement
20VC (20 Minute VC)
Aug 3
20VC: 70% of Neolabs Will Die | There Will be a $100BN US Open-Source Model | Data is a Trillion $ Market | Governments Cannot Regulate Models: It is Too Late | The Cyber Attacks to Come Will be Insane with Anastasios Angelopoulos @ Arena
Investing for Beginners
Jul 27
The Stoplight System with Tykr founder Sean Tepper
a16z Podcast
Jul 17
Amjad Masad on Going Direct, Building Replit, and the Future of Software
The Vergecast
Jul 16
The one AI detector people actually trust
Explore Related Topics
This podcast is featured in Best Finance Podcasts (2026) — ranked and reviewed with AI summaries.
Read this week's Startups & Product Podcast Insights — cross-podcast analysis updated weekly.
You're clearly into Odd Lots.
Every Monday, we deliver AI summaries of the latest episodes from Odd Lots and 192+ other podcasts. Free for one show.
Start My Monday DigestNo credit card · Unsubscribe anytime