
Anthropic released Claude Opus 5, described as a thoughtful, proactive model approaching frontier intelligence at about half the price of Fable, and the video reviews the announcement, benchmarks, and demos. Opus 5 is shown outperforming Fable on several benchmarks (including agentic terminal coding, knowledge work, agentic search, computer use, business workflows, agentic coding, and biology), though Fable leads in areas like legal/health and multidisciplinary reasoning. Comparisons to GPT 5.6 Sol show Opus leading on stated benchmarks, with emphasis on cost-per-task charts across effort levels for computer use and automation workflows. The script covers ArcAGI and ARC AGI 3 performance and high costs, alignment and behavioral audit results, vulnerability/exploit capabilities and safeguards, plus CursorBench and coding-agent index tradeoffs. It notes availability via API/web/Claude Code, pricing ($5/M input, $25/M output), a faster mode at 2x price, and new features like mid-conversation tool changes and automatic safety fallbacks. 00:00 Opus 5 Overview 00:15 Benchmark Wins 00:54 Cost Per Task 02:04 ArcAGI Results 02:33 Safety And Alignment 03:17 Cybersecurity Tradeoffs 04:29 Coding Benchmarks 05:53 Model Demos 06:12 Reviews And Pricing 07:07 Platform Updates 07:49 Wrap Up
Technical content at the intersection of AI and development. Building with AI agents, Claude Code, and modern dev tools - then showing you exactly how it works.
Weekly deep dives on AI agents, coding tools, and building with LLMs - delivered to your inbox.
Free forever. No spam.
Subscribe Free
New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.