Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)
(news.ycombinator.com)
1.
2.
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
(news.ycombinator.com)