Hugging Face daily papers·5d agoRRSI: Regularized Recursive Self-Improvement of Agent Harnesses#rrsi#agent-harness#recursive-self-improvement 2 sources
arXiv cs.AI / cs.LG / cs.CL·10d agoDouble descent is the principle of least action#double-descent#generalization#learning-theoryAI research
arXiv cs.AI / cs.LG / cs.CL·15d agoA Unified and Constrained View of Regularization-Based Robust Reinforcement Learning#adversarial-robustness#constrained-optimization#deep-rlAI research1
arXiv cs.AI / cs.LG / cs.CL·16d agoData Scarcity and Model Sparsity: Mixtures-of-Experts Overfit More to Repeated Data#data-repetition#mixture-of-experts#overfittingAI research
arXiv cs.AI / cs.LG / cs.CL·16d agoOn the Regularization Landscape for the Linear Recommendation Models#low-rank#matrix-factorization#recommendation-systemsAI research