3 papers
cs.CV2026
SG-Layout: Structured Scene Graph-Guided Layout Generation with LLMs
Junsheng Wang, Chao Chen, Mengying Xie +2
Understanding and generating spatially coherent layouts from natural language remains a fundamental yet challenging task for large language models (LLMs). Existing LLMs often strug…
cs.DC2026
EdgeCoInfer: Hierarchical Collaborative Inference for On-Device Multimodal Large Models
Lin Tan, David K. Y. Yau, Songtao Guo +1
To deliver ubiquitous intelligence, modern mobile applications increasingly execute concurrent Multimodal Large Language Models (MLLMs) on edge devices, presenting severe challenge…
cs.RO2026
One-to-Two Acting: A Novel Framework for Single-arm Agent Action Expansion to Dual Arms
Youbin Yao, Nieqin Cao, Mingyan Li +3
Dual-arm manipulation can improve throughput via parallel execution, but collecting bimanual demonstrations for training is costly and difficult. We present ExS2D, a hierarchical a…