Claude Opus 5
Listen to episode
About this episode
Anthropic shipped Claude Opus 5 on July 24th at the same price as the model it replaces, and buried the interesting part in a footnote: turn the effort dial to max and the scores go down. Host Emily Laird reads the system card, separates the vendor-run benchmarks from the independently administered ones, and explains why extra test-time compute buys ambition rather than correctness. Also covered: three outages in two days, a cyber classifier that quietly routes part of your traffic to an older model, and why Anthropic's own coding guidance stops one rung short of the top setting. If your team is paying for maximum thinking, you may be paying for scope creep with a token bill attached.
🎯 JOIN THE AI WEEKLY MEETUPS
https://www.uwstout.edu/ai-weekly-meetup
📩 EMAIL REMINDERS FOR THE MEETUPS
https://app.e2ma.net/app2/audience/signup/2101263/1779703/
💬 CONNECT WITH EMILY LAIRD ON LINKEDIN
http://www.linkedin.com/in/meet-emily-laird
More AI podcast episodes
Browse all →Want to find AI jobs?
Join thousands of AI professionals finding their next opportunity