positiveSYS.SOURCE: ICLR 2026• 2026-09-18T18:55:35Z
Direct Semantic Communication Between Large Language Models via Cache-to-Cache Mechanism
This paper introduces Cache-to-Cache (C2C), a method enabling direct semantic communication between large language models by leveraging KV-cache projections and a learnable gating mechanism. It achieves 6.4-14.2% higher accuracy and 2.5x speedup compared to text-based communication.
*** END OF TRANSMISSION ***