1 paper
Xingyu Qu, Siyuan Lu, Zhiyu Chen +2
Sharing context between LLMs in a multi-model system requires the receiving model to prefill the shared prefix because KV caches are model-specific. Recent closed-form cross-model…