collaborators

7 papers

cs.CL2026

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

Yilong Li, Suman Banerjee, Tong Che

Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@ as the canonical metric. Yet the standard policy class draws…

cs.AR2026

MEDUSA: Scalable Biometric Sensing in the Wild through Distributed MIMO Radars

Yilong Li, Ramanujan K Sheshadri, Karthik Sundaresan +2

Radar-based techniques for detecting vital signs have shown promise for continuous contactless vital sign sensing and healthcare applications. However, real-world indoor environmen…

cs.SD2026

NPUsper: Eliminating Redundant Computation for Real-Time Whisper on Mobile NPUs

Sihyeon Lee, Hojeong Lee, Sungwon Woo +3

We present NPUsper, a live transcription system that makes Whisper efficient on mobile NPUs by eliminating redundant computation. To avoid the heavy padding used by prior streaming…

cs.DC2026

Tiny but Mighty: A Software-Hardware Co-Design Approach for Efficient Multimodal Inference on Battery-Powered Small Devices

Yilong Li, Shuai Zhang, Yijing Zeng +5

Large Multimodal Models (LMMs) are inherently modular, comprising vision and audio encoders, a projector, and a language backbone. Yet existing systems execute them monolithically,…

cs.AI2025

Babel: A Scalable Pre-trained Model for Multi-Modal Sensing via Expandable Modality Alignment

Shenghong Dai, Shiqi Jiang, Yifan Yang +4

This paper presents Babel, the expandable modality alignment model, specially designed for multi-modal sensing. While there has been considerable work on multi-modality alignment,…

cs.AI2025

AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection

Weidi Luo, Shenghong Dai, Xiaogeng Liu +4

The rapid advancements in Large Language Models (LLMs) have enabled their deployment as autonomous agents for handling complex tasks in dynamic environments. These LLMs demonstrate…