Hugging Face daily papers·3d agoViRDM: Taming Representation Distribution Matching for Few-Step Causal Video Generation#virdm#video-generation#diffusion
Hugging Face trending models·3d agoViggle/Qwen-Image-2.1-viggle-turbo — new model trending #30 on Hugging Face#qwen#qwen-image#viggleModel release 9 sources 15 min
Hugging Face daily papers·4d agoSix Layers Less: Encoder Pruning for Whisper with Label-Free Recovery#whisper#asr#pruning
arXiv cs.AI / cs.LG / cs.CL·4d agoTrain Where the Quantized Model Goes: On-Policy Distillation for Low-Bit Reasoning#quantization#distillation#on-policyAI research
Interconnects·4d agoDebating RSI, the US-China Gap, and Jaggedness with JS Denain of Epoch AI#epoch-ai#rsi#us-chinaAI research 15 min
The Decoder·4d agoXiaomi's affordable flagship AI leads the open models, and Anthropic says Claude helped get it there#xiaomi#mimo-v2.6#open-weights 5 sources 3 min
arXiv cs.AI / cs.LG / cs.CL·5d agoHarness-Zero: Harness Distillation via Agent-as-Harness#agents#harness#distillationAI research 2 sources
Hugging Face daily papers·6d agoACLArena: Agent Continue Learning in Multi-stage Post-training#continual-learning#agents#lora
Hugging Face daily papers·6d agoThink Like a World Model, Act Like a VLA: Distilling World-Model Representations into Compact Robot Policies#vla#world-model#robotics
Hugging Face daily papers·7d agoOne to More, More to One: Category-Aware Iterative Expert Training for Software Engineering Agents#swe-agents#reinforcement-learning#distillation1
Hugging Face daily papers·8d agoSpatial-Interactor: Learning Spatial Reasoning through Interaction with the Observable Physical World#spatial-reasoning#vision-language-models#curriculum-learning
Hugging Face daily papers·9d agoCalibrating Teacher--Student Discrepancy for On-Policy Distillation#distillation#knowledge-distillation#on-policy-distillation
arXiv cs.AI / cs.LG / cs.CL·9d agoRetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning#distillation#llm-agents#multi-turn-agentsAI research
arXiv cs.CR·9d agoFingerprinting Multimodal Large Language Models#attention#distillation#fingerprintingAI safety & security
arXiv cs.AI / cs.LG / cs.CL·9d agoMulti-Dimensional Prosody Judgment For Live Streaming Speech Synthesis#distillation#grpo#llm-judgeAI research1
Hugging Face daily papers·10d agoWhen EOS Tokens Disagree: Understanding Length Inflation in On-Policy Distillation#distillation#eos-tokens#gemma
Hugging Face daily papers·10d agoVideo DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation#diffusion-transformer#distillation#inference-efficiency
Hugging Face daily papers·10d agoRetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning#agentic-rl#alfworld#distillation
arXiv cs.AI / cs.LG / cs.CL·11d agoCoupled Calibration and Learning: Mitigating Teacher Bias in LLM Distillation without Target-Domain Reward Feedback#covariate-shift#distillation#knowledge-transferAI research1
CSO Online·11d ago highThreat actors are coming for your AI assets to operationalize their use of AI#ai#api-keys#apt42 in the wild 5 min
arXiv cs.AI / cs.LG / cs.CL·12d agoMind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States#distillation#llm#personalizationAI research1
Hugging Face daily papers·13d agoMind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States#distillation#human-aware-ai#llm-training1
TechCrunch · AI·15d agoY Combinator’s Garry Tan wants U.S. open-weight AI labs to ‘distill’ frontier models, too#ai-policy#anthropic#dario-amodei 2 min