Claude’s new model is more ‘honest’ when it messes up
Anthropic is releasing Claude Opus 4.8 on Thursday, and the company is touting the model's "honesty." According to Anthropic, it trains "all [its] models to be honest - for instance, to avoid making claims that they can't support." But it notes that "a general problem with AI models is that they sometimes jump to conclusions, […]
Original Source
Read the full article at Theverge →KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.