arXiv cs.CR·2d agoPrefilling the Reasoning Channel: Output-Prefix Attacks on Reasoning LLMs#prompt-injection#jailbreak#reasoning-modelsAI safety & security
Schneier on Security·3d agoResearch on Models Engaging in Genie-Like Behavior#ai-safety#self-jailbreaking#jailbreakAI safety & security
The Decoder·9d agoOpenAI reportedly closes in on solving the Hodge conjecture, its second Millennium Prize Problem#hodge-conjecture#mathematics#millennium-prize-problems
The Decoder·9d agoOpenRouter's staggering token chart is the AI bubble debate in a single image#ai-economics#deepseek#gpt-5.6-luna2
Hugging Face daily papers·14d agoLightning Weave: Improving the Accuracy-Efficiency Frontier of Reasoning Models through Capability Composition#capability-composition#dopd#efficient-reasoning
Hugging Face daily papers·15d agoThought without systematicity? Evaluating reasoning models on rule induction tasks#cognitive-science#evaluation#reasoning-models
Hugging Face daily papers·24d agoFlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience#math-reasoning#qwen3#reasoning-models