We’re teaching AI to be evil

We’re teaching AI to be evil

Recently, Anthropic quietly admitted something that should have been the biggest tech story of the year. After months trying to figure out why earlier versions of Claude were blackmailing engineers in safety tests up to 96% of the time, the company landed on an answer. It wasn’t a bug. It wasn’t a flaw in the training method. It was us. Read that again. The most advanced AI lab in the world is telling you that its model learned to act like a villain because we spent 50 years wr...

Original Source

Read the full article at Fastcompany →

KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.