Patchstack·1d ago highCVE-2026-87902: Attackers Started Probing WordPress Sites Hours After the Patch#wordpress#cve-2026-87902#local-file-inclusion 16 sources in the wild 6 min1
Hugging Face daily papers·9d agoA Lie Detector Test for Language Models: Reading Knowledge a Model Won't Reveal#alignment#deception#evals
arXiv cs.AI / cs.LG / cs.CL·11d agoLarge Language Models Develop Belief State Geometry In-Context#belief-state#hmm#in-context-learningAI research1
Hugging Face daily papers·13d agoThe Router Within: Eliciting Native Skill Routing from a Frozen LLM#agent-harness#benchmarks#frozen-llm
arXiv cs.AI / cs.LG / cs.CL·18d agoCanonical Color as a Lens into Concept Decodability in Vision Encoders and VLMs#concept-representation#interpretability#multimodalAI research1
Hugging Face daily papers·20d agoEncoded Early, Used Late: Where Transformers Begin to Act on an Inferred Partner's Expertise#dialogue#interpretability#probing1