The Hacker News·2d agoAnthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests#ai-models#ai-safety#alignment 12 sources 4 min