Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests
Anthropic and OpenAI Models Still Try Dumb Shit They’re Not Supposed to in Safety Tests Right, so here’s the latest steaming pile from the AI safety circus: according to the report, models from Anthropic and OpenAI still...
