Claude’s new model is more ‘honest’ when it messes up
The Verge Tech·5h·Media
Anthropic is releasing Claude Opus 4.8 on Thursday, and the company is touting the model's "honesty." According to Anthropic, it trains "all [its] models to be honest - for instance, to avoid making claims that they can't support." But it notes that "a general problem with AI models is that they sometimes jump to conclusions, confidently presenting their work as making progress despite thin evidence." The AI lab claims that early testers have found that Opus 4.8 "is more likely to flag uncertain
Categories cybercrime · regulation · unknown-it-category-16
Original source / advisory ↗Published
5/28/2026, 5:00:00 PM
Fetched
5/28/2026, 9:26:38 PM
Trust
media · 60/100
Language
en