arXiv cs.AI / cs.LG / cs.CL·9d agoLocal Sparsity Enables Unsupervised LLM Safety Detection#activation-analysis#anomaly-detection#linear-representation-hypothesisAI safety & security