arXiv cs.AI / cs.LG / cs.CL·3d agoNonequilibrium Phases of Repulsive Self-Attention: Chaos, Attention Condensation, and Emergent Locality#transformers#self-attention#dynamical-systemsAI research
Hugging Face Blog·5d agoTransformers now runs llama.cpp quants#transformers#llama.cpp#quantizationAI tools & infra
Hugging Face daily papers·5d agoGeoPair: Geometry-Preserving Cross-Layer Factorization for Training-Free Transformer Compression#model-compression#transformers#factorization1
arXiv cs.CR·9d agoOrigin Is All You Need: Provenance-Aware Transformers for Structural Trust-Boundary Separation#indirect-prompt-injection#provenance#transformersAI safety & security
arXiv cs.AI / cs.LG / cs.CL·10d agoHow Model Growth, Recursion, and Boundary Operators Influence Scaling Exponents#compute-efficiency#looped-transformers#model-growthAI research
arXiv cs.AI / cs.LG / cs.CL·12d agoDisentangling Representation Evolution in Transformers through Directional Decomposition#interpretability#model-editing#pretrainingAI research1
Hugging Face daily papers·13d agoDisentangling Representation Evolution in Transformers through Directional Decomposition#compression#interpretability#model-editing
Hacker News · AI·15d agoRetrospectively Reverse-Engineering Apple's Neural Engine#apple#hardware-architecture#m1AI research 15 min2
arXiv cs.AI / cs.LG / cs.CL·15d agoType Diversity Enables Transformers to Generalise Compositionally#cogs#compositional-generalization#grammatical-frameworkAI research
arXiv cs.AI / cs.LG / cs.CL·15d agoSAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking#attention-sparsification#efficiency#flashattentionAI research2
Hugging Face daily papers·16d agoSAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking#attention#efficiency#long-context1
arXiv cs.AI / cs.LG / cs.CL·16d agoDistance generalization in transformers: why bother with positional encoding?#alibi#length-generalization#nopeAI research2
arXiv cs.AI / cs.LG / cs.CL·16d agoBiology-in-the-loop: Amortized Adaptive Hit Discovery in CRISPR Screens#adaptive-experimentation#assayloop#benchmarkAI research1
arXiv cs.AI / cs.LG / cs.CL·18d agoIt's Not RoPE that Creates Sinks: The Role of Self-Concentration and Value-Non-Mixing in Attention#attention-sinks#interpretability#llmAI research
Hugging Face daily papers·20d agoEncoded Early, Used Late: Where Transformers Begin to Act on an Inferred Partner's Expertise#dialogue#interpretability#probing1
CERT/CC Vulnerability Notes·25d agoVU#456290: Hugging Face Transformers library writes remote code to disk prior to consent check#cve-2026-80047#hugging-face#machine-learningVulnerability1