Hacker News · AI·8d agoCache-to-Cache: Direct Semantic Communication Between LLMs (2025)#arxiv#inference#kv-cacheAI research1