arXiv cs.AI / cs.LG / cs.CL·12d agoCorrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection#ai-safety#chain-of-thought-monitoring#deceptionAI safety & security
The Decoder·19d agoOpenAI reports AI "research interns" and warns about its own pace at the same time#agentic-ai#alignment#automated-research 5 min