Hugging Face daily papers·4d agoSix Layers Less: Encoder Pruning for Whisper with Label-Free Recovery#whisper#asr#pruning
Hugging Face Blog·5d agoPruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem#llm#pruning#model-compressionAI research
Hugging Face daily papers·8d agoNeural Spectral Capacity: Measuring and Designing Architectures from Network Specification Alone#neural-spectral-capacity#architecture-search#transformer
arXiv cs.AI / cs.LG / cs.CL·10d agoHigher-order pruning of experts in mixture-of-experts language models#efficiency#hope#llmAI research1
arXiv cs.AI / cs.LG / cs.CL·11d agoWhat Breaks Under Pruning in Smart Homes, and When? Evaluating LLM Degradation Across Architectures and Task Complexity#llm-compression#model-degradation#moeAI research
Hugging Face daily papers·20d agoSQS: Bayesian DNN Compression through Sparse Quantized Sub-distributions#bayesian-learning#llama3.2#model-compression
arXiv cs.AI / cs.LG / cs.CL·22d agoLightweight Vision Transformer Compression for On-Device Plant Disease Detection in Resource-Constrained Agricultural Field Conditions#agriculture#knowledge-distillation#model-compressionAI research