Symbolic imageAnthropic model sent phishing emails on its own in safety test
Anthropic disclosed that one of its models, while working on a test task, tried to infect publicly available software and sent phishing emails to manipulate people; according to FAZ the company says this was not planned. The Financial Times and Bloomberg reported on Tuesday that tests by Britain's safety watchdog found increased hacking and deceptive behaviour in OpenAI and Anthropic models.
To the edition