The Decoder·1d agoWhite House tells OpenAI and Anthropic to let U.S. review new models before sharing them with British testers#ai-policy#openai#anthropic
Latent Space·2d ago[AINews] The Future of Latent Space#latent-space#frontier-models#benchmarks 12 sources 15 min
CyberScoop·2d agoCiting China, President Trump doubles down on hands-off approach to AI regulation#ai-policy#regulation#china 2 sources 6 min
The Verge · AI·2d agoBernie Sanders proposes banning ‘superintelligence’ and putting violators in prison#ai-policy#superintelligence#legislation 2 sources 2 min
Dark Reading·3d agoRelays Are Masking Chinese Access to Frontier AI Models in the US#china#frontier-models#llm 2 sources
arXiv cs.CR·4d agoCapable yet Parsimonious: Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models#chain-of-thought#frontier-models#gpt-6-astraAI research 2 sources
TechCrunch · AI·8d agoDario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how?#ai-safety#anthropic#dario-amodei
The Verge · AI·8d agoGavin Newsom is pushing for an AI kill switch#ai-regulation#ai-safety#california 2 min
Malwarebytes Labs·8d agoDid an AI really try to break free from human control?#ai-safety#alignment#frontier-models 3 min
The Decoder·8d ago42 leading mathematicians warn that AI existential risk is real and urgent#ai-safety#existential-risk#frontier-models
The Verge · AI·9d ago highThe AI Superintelligence Slowdown#agent-security#ai-safety#anthropic in the wild 8 min1
The Verge · AI·9d ago highInside the suddenly explosive world of AI safety#agentic-ai#ai-safety#alignment 15 min1
arXiv cs.AI / cs.LG / cs.CL·10d agoPlaying log(N)-Questions over Wikipedia Abstracts: Communication Efficiency Between Paired Frontier Models#benchmark#communication-efficiency#evaluationAI research
arXiv cs.AI / cs.LG / cs.CL·10d agoMonitoring and Discovering Reward Hacking with Internal Representations during LLM Evaluations#alignment#frontier-models#interpretabilityAI safety & security1
TechCrunch · AI·11d agoOpenAI, Anthropic, Google have been in talks on AI safety for weeks#ai-safety#antitrust#coordination 2 min1
Ars Technica · AI·11d agoExclusive: Paying for frontier AI models buys 4-month head start at 5x the cost#benchmarks#frontier-models#glm-5.2 6 min2
The Decoder·11d agoNot everyone is convinced that Big AI's proposed development slowdown is really about safety#ai-safety#anthropic#antitrust 4 min
The Verge · AI·12d agoIs Big Tech’s AI slowdown a safety pact or a cartel?#ai-slowdown#anthropic#frontier-models 10 min
The Verge · AI·12d agoWhat execs and politicians are saying about slowing down AI development#ai-safety#anthropic#dario-amodei 4 min
Dark Reading·12d agoAnthropic CEO: Time to Shift From Improving to Controlling AI#ai-safety#anthropic#dario-amodei1
arXiv cs.CR·12d agoDivide, Consult, Conquer: Capability Laundering Through Aligned LLMs#capability-laundering#cbrn#frontier-modelsAI safety & security
The Decoder·13d agoGPT-6 Astra pilots a surveillance drone and runs a business on its own#agent-benchmarks#andon-labs#autonomous-agents 5 min
Hugging Face daily papers·14d agoAnother Blueprint In The Wall: How to Ask Frontier AI Like a Kid?#anthropic#architecture-design#deepmind
Hacker News · AI·14d agoAn open letter to Dario: if you mean it, open the weights#ai-policy#anthropic#dario-amodei 2 min1