Claude
Anthropic's AI assistant with advanced conversation, document analysis, and a massive context window.
About Claude
Tags
Reviews (3)
Claude Opus 5 Review: The Smartest Model Anthropic Has Shipped — and the Most Annoying
Opus 5 tops the Artificial Analysis Intelligence Index at 61 and posted a near-4x jump on ARC-AGI-3, all at the same price as Opus 4.8. Then the reviewers who tested it called it 'brilliant but annoying': it argues with instructions, stops early, does unrequested refactors, and breaks prompt scaffolding tuned for Opus 4.8. One more finding deserves attention — a reported jump in its hallucination rate.
- +Ranked first on the Artificial Analysis Intelligence Index at 61
- +ARC-AGI-3 leap to 30.2% from a 7.8% prior best
- +Within 0.5% of Fable 5 on CursorBench at half the cost
- -Argues with instructions and stops before finishing work
- -Reported hallucination rate up 14 points to 50%
- -Verbose by default — confirmed in Anthropic's own docs
Claude Fable 5 Review: The Most Capable Public Model Yet — for a Premium
Anthropic's Claude Fable 5 is the most capable model the public can use today — topping SWE-Bench Pro and excelling at vision and long tasks. It is also the priciest major model and ships with hard safety guardrails. Our first-look verdict.
- +Best agentic-coding result to date (80.3 on SWE-Bench Pro)
- +State-of-the-art vision plus stronger long-horizon autonomy
- +Mythos-class capability finally available to everyone
- -Most expensive major model at $10 in / $50 out per million tokens
- -Safety classifiers can block legitimate security research
- -Mandatory 30-day retention even on zero-retention plans
Claude Review (July 2026): Fable 5, Opus 4.8 & the New Sonnet 5
Anthropic's lineup leveled up: Fable 5 is the most capable public model, Opus 4.8 runs hundreds of parallel subagents, and Sonnet 5 is the new Pro default at a fraction of the cost. Where Claude leads, where it still struggles, and whether Pro/Max are worth it in mid-2026.
- +Best-in-class nuanced writing
- +Fable 5 tops SWE-Bench Pro (80.3) for public coding
- +Opus 4.8 Dynamic Workflows run hundreds of parallel subagents
- -Fable 5 is the priciest major model ($10/$50 per M)
- -Hard safety guardrails block some legitimate cyber/bio work
- -Still no native image generation
Similar Tools
ChatGPT
Text & ChatOpenAI's GPT-4o powered chatbot. Conversation, writing, image analysis, coding and more.
Perplexity
Text & ChatAI-powered search engine that answers questions with cited sources. A smarter alternative to Google.
Gemini
Text & ChatGoogle's AI model with live internet access, image analysis, and deep Google service integration.