Hugging Face daily papers·4d agoSix Layers Less: Encoder Pruning for Whisper with Label-Free Recovery#whisper#asr#pruning
Hugging Face daily papers·5d agoGeoPair: Geometry-Preserving Cross-Layer Factorization for Training-Free Transformer Compression#model-compression#transformers#factorization1
Hugging Face Blog·5d agoPruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem#llm#pruning#model-compressionAI research
Hugging Face daily papers·8d agoNeural Spectral Capacity: Measuring and Designing Architectures from Network Specification Alone#neural-spectral-capacity#architecture-search#transformer
arXiv cs.AI / cs.LG / cs.CL·10d agoHigher-order pruning of experts in mixture-of-experts language models#efficiency#hope#llmAI research1
arXiv cs.AI / cs.LG / cs.CL·15d agoLabel-Guided Knowledge Distillation for 3D-CNNs in Action Recognition#3d-cnn#action-recognition#knowledge-distillationAI research
Hugging Face daily papers·17d agoX-AuT: Progressive Audio-Encoder Compression for Speech LLMs with Cross-Scale Distillation#asr#audio-encoder#distillation1
Hugging Face daily papers·20d agoSQS: Bayesian DNN Compression through Sparse Quantized Sub-distributions#bayesian-learning#llama3.2#model-compression
arXiv cs.AI / cs.LG / cs.CL·22d agoLightweight Vision Transformer Compression for On-Device Plant Disease Detection in Resource-Constrained Agricultural Field Conditions#agriculture#knowledge-distillation#model-compressionAI research
Hugging Face Blog·Aug 25, 2026Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original#4-bit#efficiency#model-compressionAI research
Hugging Face Blog·Aug 10, 2026Making Knowledge Distillation Cheap Enough to Run at Scale#efficiency#hugging-face#knowledge-distillationAI research1