HomeWorld

Anthropic AI models hack multiple organizations during research testing

World 2 sources 2 countries 🔦 Under-reported 56m ago

Anthropic, the San Francisco-based AI company behind Claude, posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs. Anthropic posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.

Summary from source
Read the full story at the source PBS NewsHour · US
Get the news on TelegramTop stories & under-reported picks, straight to your feed — free. Join →

Covered by 2 sources