Nov 26, 2025 · 59m · startup-ideas
Reviewing Claude Opus 4.5
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
Greg Isenberg and James (The Boring Marketer) run a live, head-to-head coding benchmark comparing Anthropic's Claude Opus 4.5 against Google's Gemini 3 Pro by building a full-stack SaaS application from scratch. Through hands-on landing page generation, interactive prototyping, and an analysis of custom Claude skills, they explore the evolving frontier of non-technical software development.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Greg holds 22.9% of the talking time here. How this is scored →
speaking balance: gold is Greg, purple is the guest (3 minute bins)
James gently corrects Greg's assumption that custom copywriting skills generated the copy, noting only the design skill was active.
Hardest push from Greg ▶ 25:59 Host delivers direct negative critique of ad designWhen asked for an honest evaluation of James's ad, Greg directly refuses to flatter it, pointing out that it is too busy and has too much text.
Biggest teaching moment ▶ 46:00 Systematic breakdown of Claude skill developmentJames provides a comprehensive masterclass on configuring custom Claude skills using Perplexity MCP research on industry leaders combined with brand voice distillation.
Greg holds their own ▶ 55:20 Host frames content as the foundational UX elementGreg demonstrates product expertise by reframing why vibe-coded apps fail to convert, explaining that words are the actual UX.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | Greg as informed peer | Guest teaching | Guest disagreement | Greg pushing back | Why |
|---|---|---|---|---|---|---|
| Episode Overview and Experiment Setup | 4 | 2 | 0 | 0 | Greg sets the experiment parameters by selecting a startup concept (EstateClear) from IdeaBrowser, while James sets up the side-by-side prompt testing between Claude Opus 4.5 and Gemini 3 Pro. The dynamic is fully cooperative and exploratory. | |
| Sponsor: IdeaBrowser.com Promotion | 4 | 2 | 0 | 0 | Following Greg's IdeaBrowser sponsor read, James explains why he runs the test in Google AI Studio to isolate variables, and both make friendly predictions about model behaviors. | |
| Reviewing Claude Opus 4.5's Landing Page Output | 3 | 3 | 1 | 0 | Greg asks whether custom copy skills generated the landing page text, and James clarifies that only the front-end design skill was used, explaining that Opus produced the copy natively. | |
| Comparing Gemini 3 Pro's Landing Page Build | 5 | 2 | 0 | 1 | Greg points out that Gemini included an AI copywriting widget just as he had predicted, evaluating its layout and visual aesthetics against Opus 4.5. | |
| Prompting Full Clickable Prototypes and Backend Architecture | 3 | 5 | 0 | 0 | Greg inquires about building full SaaS backends, prompting James to explain his production stack using Neon, Clerk, Vercel, and Stripe, comparing Claude's end-to-end execution to other models. | |
| Analyzing Google's Vertically Integrated Ecosystem | 5 | 4 | 0 | 0 | James reviews Google's vertical integration across hardware TPUs, workspace data, and developer tools, while Greg adds business context regarding Google's equity stake in Anthropic. | |
| Exploring Nano Banana Pro and AI Ad Generation | 6 | 3 | 0 | 5 | James shows his Meta ad creatives generated with Nano Banana Pro and asks for honest feedback. Greg delivers candid pushback, critiquing the ad as overly busy and text-heavy before highlighting its workable visual elements. | |
| Evaluating Claude Opus 4.5's Interactive App Prototype | 6 | 4 | 1 | 1 | James questions the recurring use case of an estate probate tool, and Greg elaborates on vertical SaaS mechanics, retention hurdles, and customer discovery strategies. | |
| Reviewing Gemini Prototype and Debugging Strategies | 3 | 6 | 0 | 0 | When Gemini generates console errors, James educates Greg on his prompt debugging framework, including using terminal 'ultra think' commands and prompting specialized sub-agents to trace root causes. | |
| Testing Google Anti-gravity IDE and Browser Integration | 5 | 5 | 0 | 2 | Greg wagers that Google's Anti-gravity IDE will yield a worse design output, while James explains the value of Anti-gravity's Chrome extension for direct DOM inspection during debugging. | |
| Anti-gravity Output Review and Workflow Insights | 6 | 2 | 0 | 0 | The test validates Greg's design prediction, but both note Anti-gravity's autonomous use of Nano Banana to create mockups first, extracting UI wireframe lessons. | |
| Mastering Claude Skills and Elevated Direct Response | 3 | 7 | 0 | 0 | James delivers an in-depth walkthrough of building custom Claude skills for 'elevated direct response,' combining automated research via Perplexity MCP with personal writing voice distillation. | |
| The Claude Front-End Design Skill and Copy Optimization | 6 | 5 | 0 | 0 | James demonstrates Anthropic's official front-end design plugin, and Greg synthesizes the discussion by asserting that content is UX and copy drives vibe coding conversion rates. |