The Hacker News·2d agoAnthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests#ai-models#ai-safety#alignment 12 sources 4 min
TechCrunch · AI·10d agoAnthropic and OpenAI want to embed safety evaluators. Will they really be independent?#alignment#anthropic#apollo-research 7 min