Skip to main content
The AI Breakdown

The Most Useful New AI Features and Tools to Try

34 min episode · 2 min read

Episode

34 min

Read time

2 min

Topics

Fundraising & VC, Artificial Intelligence, Software Development

AI-Generated Summary

Key Takeaways

  • Claude Browser Integration: Claude's desktop app now includes a built-in browser that operates independently from your personal browsers and logins, enabling agentic web tasks like form-filling and web app interaction without additional installation. This shifts Claude from answering questions about web tasks to directly executing them, narrowing the gap between AI assistant and AI operator.
  • ChatGPT Temporary Chat Upgrade: ChatGPT now allows users to selectively pull existing memories, custom instructions, and plugins into temporary chats without permanently saving the session. Users can also retroactively save a temporary chat to history. This gives privacy-conscious users a practical middle path between full memory persistence and completely isolated sessions.
  • Gemini 3.5 Transcribe Smart Mode: Google's new speech-to-text model offers a "smart transcribe" mode that strips filler words, condenses rambling, and resolves self-corrections to produce clean, intent-accurate text across 85-plus languages. It supports up to three speaker distinctions, custom vocabulary for jargon, and functions in noisy environments, making it viable infrastructure for third-party voice app developers.
  • Pflow H3 Max Speed Benchmark: Pflow's H3 Max video model generates clips in an average of 3.49 seconds, compared to 26.7 seconds for Gemini Omni Flash and 5.68 seconds for Seed Dance 2.5. At roughly 20 cents per generation, this sub-real-time output speed opens practical use cases for live interactive visual environments and responsive multiplayer generated worlds.
  • Multiple Gmail Accounts in ChatGPT Plugins: ChatGPT now supports connecting multiple Gmail accounts simultaneously within its plugins, removing a long-standing limitation that forced power users to choose a single connected inbox. For professionals managing several email addresses across clients or projects, this single update meaningfully expands the practical utility of ChatGPT's agentic email capabilities.

What It Covers

This episode catalogs a dozen new AI features released within a single week, spanning Claude's built-in browser for desktop, ChatGPT Work's cloud computer, Gemini 3.5 Transcribe's intent-cleaning voice model, Google's Veo Omni 1.1 Flash video generator, and Pflow's H3 Max video model generating clips faster than real-time playback.

Key Questions Answered

  • Claude Browser Integration: Claude's desktop app now includes a built-in browser that operates independently from your personal browsers and logins, enabling agentic web tasks like form-filling and web app interaction without additional installation. This shifts Claude from answering questions about web tasks to directly executing them, narrowing the gap between AI assistant and AI operator.
  • ChatGPT Temporary Chat Upgrade: ChatGPT now allows users to selectively pull existing memories, custom instructions, and plugins into temporary chats without permanently saving the session. Users can also retroactively save a temporary chat to history. This gives privacy-conscious users a practical middle path between full memory persistence and completely isolated sessions.
  • Gemini 3.5 Transcribe Smart Mode: Google's new speech-to-text model offers a "smart transcribe" mode that strips filler words, condenses rambling, and resolves self-corrections to produce clean, intent-accurate text across 85-plus languages. It supports up to three speaker distinctions, custom vocabulary for jargon, and functions in noisy environments, making it viable infrastructure for third-party voice app developers.
  • Pflow H3 Max Speed Benchmark: Pflow's H3 Max video model generates clips in an average of 3.49 seconds, compared to 26.7 seconds for Gemini Omni Flash and 5.68 seconds for Seed Dance 2.5. At roughly 20 cents per generation, this sub-real-time output speed opens practical use cases for live interactive visual environments and responsive multiplayer generated worlds.
  • Multiple Gmail Accounts in ChatGPT Plugins: ChatGPT now supports connecting multiple Gmail accounts simultaneously within its plugins, removing a long-standing limitation that forced power users to choose a single connected inbox. For professionals managing several email addresses across clients or projects, this single update meaningfully expands the practical utility of ChatGPT's agentic email capabilities.

Notable Moment

Ethan Mollick noted that H3 Max crossed a threshold where AI video generation now takes less time than watching the finished clip. This inversion — output speed exceeding playback duration — signals a structural shift in how video content could be created and consumed at scale.

Know someone who'd find this useful?

Episode Transcript

Which of these sound most exciting to you? A new ability for Claude to use a browser window to be able to do tasks for you? The ability for ChatGPT to effectively switch in and out of temporary mode as suits you? A new voice model that not only can faithfully transcribe what you actually say, but even clean it up and get at your intent, a new video model from Google with massively more controllability for actually generating videos that you can use for real things, or a totally different video model that takes you less time to generate the video than it does to watch it. These are just some of the new features, models, and tools that released this week. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. Alright, friends. Quick announcements before we dive in. First of all, thank you to today's sponsors, KPMG, Harbor, Blitsy, and HyperAgent. To get an ad free version of the show, go to patreon.com/aidailybrief, or you can subscribe on Apple Podcasts. To learn more about sponsoring the show, send us a note at sponsors@aidailybrief.ai. And lastly, if you listen to today's episode and think to yourself, man, I wanna use all those features, but I feel behind. I need to catch up in my AI skills. The next cohort of super intelligent training programs, which are incubated over here at AIDB, are coming up in September. You can find a link to them on the AI Daily Brief website, or you can just go to training.bsuper.ai. Kicking off, confirmation that Hugging Face is being sold and that, yes, it is to NVIDIA. The information broke the story late on Wednesday night following the earnings call with their sources saying that NVIDIA has agreed to buy Hugging Face for $12,900,000,000. Now rumors of the deal only started emerging over the weekend. And while, of course, this could have all been going on before anyone externally knew, it seems like it might have been a pretty speedy negotiation. Much of the reporting focused on the strong valuation multiple. Hugging Face only has a 150,000,000 in annualized revenue, meaning that Nvidia is technically paying around 80 times revenue as a price. As a point of reference, SpaceX's $60,000,000,000 acquisition of Cursor was around 15 times revenue. Still, as I said the other day when I talked about the potential of this deal, NVIDIA is very much not buying Hugging Face for their revenue. Instead, the deal seems to confirm that NVIDIA is acquiring their way into a full stack open business model. Hugging Face is the premier distribution channel for open models. NVIDIA's $6,000,000,000 licensing deal with Poolside announced last week allowed them to hire about a 100 veteran AI researchers to staff up their team, and other investments like their strategic partnership with Perplexity give them solid harnesses as well as an app player play. What's more, over the course of …

Get the full transcript (6,891 words) + summary by email — free

One-time email with the complete transcript and AI summary of this episode. No account needed.

One email, no spam. We’ll also show you what SignalCast does.

Browse all The AI Breakdown transcripts →

You just read a 3-minute summary of a 31-minute episode.

Get The AI Breakdown summarized like this every Monday — plus up to 2 more podcasts, free.

Pick Your Podcasts — Free

Keep Reading

Books, tools, and gear mentioned in this episode

SignalCast may earn commission on purchases via these links.

Tools

  • ClaudeRecommended

    by Anthropic

    Claude's desktop app now includes a built-in browser that operates independently from your personal browsers and logins, enabling agentic web tasks like form-filling and web app interaction without additional installation.
  • ChatGPTRecommended

    by OpenAI

    ChatGPT now allows users to selectively pull existing memories, custom instructions, and plugins into temporary chats without permanently saving the session.
  • by Google

    Google's new speech-to-text model offers a 'smart transcribe' mode that strips filler words, condenses rambling, and resolves self-corrections to produce clean, intent-accurate text across 85-plus languages.
  • by Google

    Google's Veo Omni 1.1 Flash video generator
  • Pflow H3 MaxRecommended

    by Pflow

    Pflow's H3 Max video model generates clips in an average of 3.49 seconds, compared to 26.7 seconds for Gemini Omni Flash and 5.68 seconds for Seed Dance 2.5.
  • Pflow's H3 Max video model generates clips in an average of 3.49 seconds, compared to 26.7 seconds for Gemini Omni Flash and 5.68 seconds for Seed Dance 2.5.
  • SPONSORS: Blitzy
  • SPONSORS: HyperAgent

company

More from The AI Breakdown

We summarize every new episode. Want them in your inbox?

Similar Episodes

Related episodes from other podcasts

Explore Related Topics

This podcast is featured in Best AI Podcasts (2026) — ranked and reviewed with AI summaries.

Read this week's AI & Machine Learning Podcast Insights — cross-podcast analysis updated weekly.

You're clearly into The AI Breakdown.

Every Monday, we deliver AI summaries of the latest episodes from The AI Breakdown and 192+ other podcasts. Free for one show.

Start My Monday Digest

No credit card · Unsubscribe anytime