Researchers from Tsinghua University and other Chinese institutions proposed Cache-to-Cache (C2C), letting multi-LLM systems exchange data via key-value caches instead of text. Experiments showed 3.1 to 5.4 percentage point accuracy gains over text handoffs and lower latency across tested model pair