arXiv cs.AI / cs.LG / cs.CL·11d agoScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents#agent-harness#ai-agents#continual-learningAI research
Hugging Face daily papers·12d agoScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents#agent-harness#continual-learning#reinforcement-learning
arXiv cs.AI / cs.LG / cs.CL·12d agoLearning to Coach for Experiential Learning#coaching#experiential-learning#llm-agentsAI research2
Hugging Face daily papers·13d agoDream-RSI: Recursive Self-Improvement through Evolving Worlds#agents#coding-agents#exploration1
Hugging Face daily papers·13d agoDiscovery Foundation Models: Toward Open-Ended Discovery Intelligence#agents#discovery-intelligence#foundation-models1
MarkTechPost·15d agoCan LLMs Engineer Their Own Agent Harness? ByteDance Seed’s HarnessDev Says Only 34 of 64 Changes Generalize#agent-harnesses#browsecomp#bytedance-seed 4 min1
Hugging Face daily papers·17d agoNegative Self-Distillation: Learning to Reason by Avoiding Flaws#distillation#llm#reasoning1
Hugging Face daily papers·19d agoProcedural Graphs: Self-Evolving Execution Structures for LLM Agents#agent-planning#llm-agents#procedural-graphs1
Hugging Face daily papers·24d agoFlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience#math-reasoning#qwen3#reasoning-models