arXiv cs.AI / cs.LG / cs.CL·18d agoCo-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails#agent-harnesses#coding-agents#fine-tuningAI research
Hugging Face daily papers·19d agoCo-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails#agentic-ai#ai-agents#fine-tuning1