works on

From the 1 of 26 linked papers with an AI index.

collaborators

26 papers

cs.DC2026

OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching

Can Xiao, Sukmin Cho, Junbong We +7

Large language model (LLM) inference serving is increasingly constrained by memory rather than compute. As long-context and long-form reasoning workloads become more prevalent, the…

cs.RO2026

Suppression Sticks, Locality Is Fragile: A Closed-Loop Target-and-Control Audit of Task-Vector Negation in VLA Policies

Shaoguang Wang, Weiyu Guo, Rushi Dai +3

Task-vector arithmetic offers a closed-form way to modify a model, yet its behavioral locality remains unclear in closed-loop robot control. We present a target-and-control audit o…

cs.RO2026

How Should Vision-Language-Action Models Use Proprioceptive State?

Yiren Zhao, Ziyang Chen, Ziyang Rao +5

Recent Vision-Language-Action (VLA) models almost universally take robot proprioceptive state as input, yet wire it in incompatible ways -- serialized into text prompts, projected…

cs.AI2026

The Geometry of Flow-Matching Uncertainty: A Cost-free Uncertainty Proxy and Its Application in Flow-based VLA Failure Detection

Ziyang Rao, Yiren Zhao, Weiyu Guo +3

The paper interprets uncertainty in flow‑matching based action models as geometric deviation in the velocity field and proposes a cost‑free proxy called denoising acceleration that…

cs.RO2026

Source-Lifted Flow Matching for Intervenable Multimodal Imitation

He Zhang, Ying Sun, Pengteng Li +6

Flow-matching policies are promising for imitation learning because they model complex multimodal action distributions. However, their stochasticity is largely passive: repeated sa…

cs.LG2026

DumpsterCluster: From Dumpster Diving to Serving LLaMA-70B on $60 GPUs

Zeyu Cao, Xuan Guo, Cheng Zhang +3

As AI datacenters retire functional GPUs, vast quantities of still capable accelerators enter secondary markets. This paper investigates whether these retired GPUs can find a produ…