Technology May 28, 2026 · 1 min read

Claude’s new model is more ‘honest’ when it messes up

Anthropic is releasing Claude Opus 4.8 on Thursday, and the company is touting the model's "honesty." According to Anthropic, it trains "all [its] models to be honest - for instance, to avoid making claims that they can't support." But it notes that "a general problem with AI models is that they som...

TH
The Verge
by Jay Peters
Claude’s new model is more ‘honest’ when it messes up
The Claude logo with a overlay of an smart phone on an orange background.

Anthropic is releasing Claude Opus 4.8 on Thursday, and the company is touting the model's "honesty."

According to Anthropic, it trains "all [its] models to be honest - for instance, to avoid making claims that they can't support." But it notes that "a general problem with AI models is that they sometimes jump to conclusions, confidently presenting their work as making progress despite thin evidence."

The AI lab claims that early testers have found that Opus 4.8 "is more likely to flag uncertainties about its work and less likely to make unsupported claims." In the company's evaluations, Opus 4.8 is "around 4x less likely than its predeces …

Read the full story at The Verge.

TH
Source

This article was originally published by The Verge and written by Jay Peters.

Read original article on The Verge
Back to Discover

Reading List