2 papers
cs.DC2026
Hestia: Hyperthread-Level Scheduling for Cloud Microservices with Interference-Aware Attention
Dingyu Yang, Fanyong Kong, Jie Dai +5
Modern cloud servers routinely co-locate multiple latency-sensitive microservice instances to improve resource efficiency. However, the diversity of microservice behaviors, coupled…
cs.DB2026
SVFusion: A CPU-GPU Co-Processing Architecture for Large-Scale Real-Time Vector Search
Yuchen Peng, Dingyu Yang, Zhongle Xie +4
Approximate Nearest Neighbor Search (ANNS) underpins modern applications such as information retrieval and recommendation. With the rapid growth of vector data, efficient indexing…