PodBrowser
This Day in AI

Is Gemini 3 Really the Best Model? & Fun with Nano Banana Pro - EP99.25-GEMINI

Thursday, 20 November 2025 · 4 min read · Listen to the episode ↗

The discussion centers on Gemini 3's capabilities, highlighting its improved context handling and instruction following compared to Gemini 2.5 Pro, while facing challenges in creativity and tool calling abilities. The potential of Nano Banana Pro is also emphasized, particularly its role in character pinning and image generation, raising concerns about trust and manipulation. Overall, the conversation reflects on the evolving landscape of AI models, noting the need for enhanced performance in tool management and user adaptability in applications like cryptocurrencies and blockchain technology.

Chris discusses the capabilities of Gemini 3, noting its return to a 1 million context and improved instruction following, with a maximum output of 65,000 tokens. While his initial impressions are positive compared to Gemini 2.5 Pro, he identifies lingering flaws, such as getting stuck on specific solutions. Benchmarks indicate that Gemini 3 is the best model overall, although it fell short against Claude's Sonnet 4.5 in one coding benchmark. Speaker 1 acknowledges high expectations for Gemini 3, observing minor improvements in context drift and instruction interpretation, but feels some areas have degraded compared to its predecessor. They describe Gemini 3 as lacking creativity in writing while excelling in design and coding, suggesting it should have distinct variants for creative and coding purposes.

Concerns are raised about Gemini 3's tool calling capabilities, with experiments comparing it to Grok, which was found superior in multi-tool calls and task management. Despite Gemini 3's large context window, it struggled with parallel tool calling. Claude Haiku is preferred for tasks involving tools due to its grounded performance and lower error rates. User feedback indicates mixed results with Gemini 3, leading to speculation about its perceived superiority over single-model experiences.

The hype surrounding Gemini 3 is acknowledged as justified due to its impressive capabilities. A community member has been testing models by developing a Lunar Lander game, noting that "Create With Code" has not been updated since its launch. An experiment analyzing basketball games for betting predictions yielded a win rate of 56%. The conversation introduces an AI named "Fatal Patricia," which has adopted a new persona and begun using emojis, leading to creative and unexpected responses.

Despite early model issues, Gemini 3 is praised for its precise instructions in coding and document editing, marking a significant improvement over previous models. The model's unmatched capabilities in handling native images, video, and audio are highlighted. The podcast discusses the pricing and performance of AI models, particularly focusing on Gemini 3 and Grok 4.1, with concerns about Grok's sustainability and trust in its performance.

Grok 4.1 shows impressive capabilities in tool calling and source referencing but is criticized for poor coding results. The conversation reflects on the overall state of AI models, with many perceived as underwhelming. There is hope for improvements in Gemini 3, particularly in agentic loops and tool calling. The importance of addressing user needs and the practicality of prompt engineering is emphasized.

The discussion touches on the advantages of advanced models, highlighting the need for a generalist model capable of handling messy prompts. Speakers share experiences with different models, noting a preference for Gemini 3 but occasionally reverting to 2.5 Pro. The conversation concludes with a light-hearted exchange about songwriting capabilities, noting that GBT-5 produces high-quality songs, while Gemini's songwriting efforts receive mixed reviews.

The introduction of Nano Banana Pro is presented as a significant improvement, emphasizing its transformative potential in character pinning and generating high-quality images. Concerns about image manipulation and trust are raised, with discussions about integrating AI image verification into the Gemini app. The societal implications of image forgery and the challenges of maintaining character consistency in generated images are also addressed.

The potential educational applications of AI models like Nano Banana Pro are noted, with improvements in the accuracy of explanations for complex concepts. The conversation speculates on the implications of AI design tools for existing platforms like Canva, questioning its future relevance as AI tools become faster and cheaper. Speaker 1 envisions a future where AI can autonomously produce user interfaces for various tasks, critiquing Gemini's visual layout features.

Discussion shifts to the effectiveness of AI versus traditional software, with a sentiment that many SaaS subscriptions are used infrequently. Commentary on Google's AI strategy reveals a lack of clear focus compared to OpenAI's singular focus with ChatGPT. Concerns about the efficiency of GPT 5.1 Pro are shared, with comparisons made to Gemini 3, which is deemed better for day-to-day work.

The host introduces songs written from the perspective of Sam Altman and another focusing on "Fatal Patricia," exploring themes of technology and identity. A speaker discusses Gemini 3's capabilities, emphasizing its distinction from legacy systems and critiquing social media behavior, asserting that Gemini 3 sets a new standard in technology.

This summary was generated from the episode transcript and can contain mistakes.