Google's NotebookLM now runs its own cloud computer with code execution and agent-based research

Google has substantially expanded NotebookLM's research capabilities by integrating Gemini 3.5 Flash, adding autonomous code execution via dedicated cloud infrastructure, and enabling agent-driven source discovery through Google Search integration. The upgrade demonstrates a shift toward agentic AI systems that can independently iterate on research tasks, with internal benchmarks showing 78% performance gains over the prior version. This positions NotebookLM as a competitive alternative to specialized research tools and signals Google's strategy to embed autonomous reasoning deeper into productivity workflows.
Modelwire context
Skeptical readThe 78% performance figure comes from Google's own internal benchmarks, with no methodology disclosed and no independent replication, which makes it nearly impossible to assess whether the number reflects genuine capability improvement or favorable eval selection.
This is largely disconnected from recent activity in our archive, as we have no prior coverage to anchor it to. It does, however, belong to a crowded competitive space where Perplexity, OpenAI's Deep Research feature, and Anthropic's Claude have all made similar 'agentic research' claims in the past year. The pattern is consistent: a productivity tool adds search-grounded iteration, calls it autonomous, and publishes a benchmark that can't be cross-checked. What's worth noting is that NotebookLM's core differentiator has historically been source-grounded summarization, and adding open-web search via Google Search integration actually muddies that proposition rather than sharpening it.
Watch whether independent researchers can reproduce the benchmark gains on a public evaluation set like GPQA or FRAMES within the next 60 days. If no external replication appears, the 78% figure should be treated as a marketing number rather than a technical one.
This analysis is generated by Modelwire’s editorial layer from our archive and the summary above. It is not a substitute for the original reporting. How we write it.
MentionsGoogle · NotebookLM · Gemini 3.5 Flash · Google Search
Modelwire Editorial
This synthesis and analysis was prepared by the Modelwire editorial team. We use advanced language models to read, ground, and connect the day’s most significant AI developments, providing original strategic context that helps practitioners and leaders stay ahead of the frontier.
Modelwire summarizes, we don’t republish. The full content lives on the-decoder.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.