Skip to content
How I AIHow I AI

Claude is BACK with Opus 5.5

I’ve been off Claude for months. Not because it got dumb, but because it got annoying. The rambling, the hedging, the preachy little disclaimers on tasks that didn’t need them. I moved most of my daily work to Codex and I didn’t miss it. Then Anthropic shipped Opus 5.5: 40% cheaper than Opus 5, faster, and with what they’re calling a fundamentally different alignment approach. I ran it for a week across real work, including four long-running agentic tasks, a full ChatPRD homepage redesign, an SVG benchmark, and one very firm refusal, and I’m ready to give you the honest verdict. There’s a lot to like. There are still two things that drive me a little crazy. And there’s one capability I genuinely wasn’t expecting. *What you’ll learn:* 1. Why I walked away from Claude entirely, and what it took for me to come back 2. The real cost math on Opus 5.5 and why pricing matters more for agentic work than single prompts 3. What happened when I ran four long-running agentic tasks, including one that tried to manipulate Claude mid-run 4. Why Opus 5.5 is now my go-to for frontend prototyping, and where it still lets me down 5. The one capability I genuinely didn’t see coming, and no other model in my stack can match it 6. The moment Opus 5.5 told me flat-out no, and what that says about where Anthropic’s safety posture actually lands in practice 7. Where Codex still wins, and how I’m splitting my model stack after a full week of testing *In this episode:* (00:00) Why I stopped using Claude (01:02) What Anthropic says Opus 5.5 is (01:54) Cost, speed, and benchmark overview (03:20) Safety, alignment, and the cybersecurity limits (05:02) How I AI bench (05:39) Voice test: is it actually not annoying? (07:54) Long-running agentic task results (10:50) Frontend prototyping (17:23) Writing voice and email (19:41) SVG illustrations (20:46) Video editing (21:42) My verdict: what it’s good at, what it still isn’t *Tools referenced:* • Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5 • ElevenLabs MCP connector: https://elevenlabs.io/mcp • Codex (OpenAI): https://openai.com/codex *Where to find Claire Vo:* ChatPRD: https://www.chatprd.ai/ Website: https://clairevo.com/ LinkedIn: https://www.linkedin.com/in/clairevo/ X: https://x.com/clairevo _Production and marketing by https://penname.co/._ _For inquiries about sponsoring the podcast, email jordan@penname.co._

Claire Vohost
Sep 22, 202624mWatch on YouTube ↗

Episode Details

EPISODE INFO

Released
September 22, 2026
Duration
24m
Channel
How I AI
Watch on YouTube
▶ Open ↗

EPISODE DESCRIPTION

I’ve been off Claude for months. Not because it got dumb, but because it got annoying. The rambling, the hedging, the preachy little disclaimers on tasks that didn’t need them. I moved most of my daily work to Codex and I didn’t miss it. Then Anthropic shipped Opus 5.5: 40% cheaper than Opus 5, faster, and with what they’re calling a fundamentally different alignment approach. I ran it for a week across real work, including four long-running agentic tasks, a full ChatPRD homepage redesign, an SVG benchmark, and one very firm refusal, and I’m ready to give you the honest verdict. There’s a lot to like. There are still two things that drive me a little crazy. And there’s one capability I genuinely wasn’t expecting. *What you’ll learn:*

  1. Why I walked away from Claude entirely, and what it took for me to come back
  2. The real cost math on Opus 5.5 and why pricing matters more for agentic work than single prompts
  3. What happened when I ran four long-running agentic tasks, including one that tried to manipulate Claude mid-run
  4. Why Opus 5.5 is now my go-to for frontend prototyping, and where it still lets me down
  5. The one capability I genuinely didn’t see coming, and no other model in my stack can match it
  6. The moment Opus 5.5 told me flat-out no, and what that says about where Anthropic’s safety posture actually lands in practice
  7. Where Codex still wins, and how I’m splitting my model stack after a full week of testing

*In this episode:* (00:00) Why I stopped using Claude (01:02) What Anthropic says Opus 5.5 is (01:54) Cost, speed, and benchmark overview (03:20) Safety, alignment, and the cybersecurity limits (05:02) How I AI bench (05:39) Voice test: is it actually not annoying? (07:54) Long-running agentic task results (10:50) Frontend prototyping (17:23) Writing voice and email (19:41) SVG illustrations (20:46) Video editing (21:42) My verdict: what it’s good at, what it still isn’t *Tools referenced:*

*Where to find Claire Vo:* ChatPRD: https://www.chatprd.ai/ Website: https://clairevo.com/ LinkedIn: https://www.linkedin.com/in/clairevo/ X: https://x.com/clairevo _Production and marketing by https://penname.co/._ _For inquiries about sponsoring the podcast, email jordan@penname.co._

SPEAKERS

  • Claire Vo

    host

    Host of the "How I AI" podcast, covering practical AI tools and workflows.

EPISODE SUMMARY

In this episode of How I AI, featuring Claire Vo, Claude is BACK with Opus 5.5 explores opus 5.5 makes Claude usable again, especially for design Claire returns to Claude after months away because Opus 5.5 finally feels minimally annoying and more concise in conversation.

RELATED EPISODES

I reviewed Opus 5.5 and GPT-6 Sol live - and the results surprised me

I reviewed Opus 5.5 and GPT-6 Sol live - and the results surprised me

Muse gets AI agent UX right

Muse gets AI agent UX right

How SpaceXAI designers use Grok Bot and Figma MCP to ship faster

How SpaceXAI designers use Grok Bot and Figma MCP to ship faster

The enterprise AI stack behind Stripe’s company brain “Kai”

The enterprise AI stack behind Stripe’s company brain “Kai”

How I use ChatGPT to run my fashion business

How I use ChatGPT to run my fashion business

I built a Claude Cowork system that does a week of PM work in a day

I built a Claude Cowork system that does a week of PM work in a day

Get more out of YouTube videos.

High quality summaries for YouTube videos. Accurate transcripts to search & find moments. Powered by ChatGPT & Claude AI.