collaborators

6 papers

cs.CV2026

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

Taiting Lu, Runze Liu, Ziwei Dong +18

Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D objects and rarely address the…

cs.CV2026

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

Taiting Lu, Kaiyuan Lin, Ziwei Dong +18

Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However, their ability to reason abou…

cs.CV2026

OmniLayout: A Schematic-Coupled Multimodal Benchmark for Constraint-Aware Geometric Reasoning in PCB Layout

Taiting Lu, Kaiyuan Lin, Mingjia Wang +12

Recent large language models (LLMs) have demonstrated remarkable progress in 3D spatial reasoning, spatial grounding, and fine-grained geometric understanding. However, their abili…

cs.CV2026

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

Taiting Lu, Kaiyuan Lin, Yuxin Tian +13

Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, their ability to convert Printed C…

cs.HC2026

PPG as a Bridge: Cross-Device Authentication for Smart Wearables with Photoplethysmography

Jiacheng Liu, Jiankai Tang, Guangye Zhao +8

As smart wearable devices become increasingly powerful and pervasive, protecting user privacy on these devices has emerged as a critical challenge. While existing authentication me…

cs.SD2024

mmWave-Whisper: Phone Call Eavesdropping and Transcription Using Millimeter-Wave Radar

Suryoday Basak, Abhijeeth Padarthi, Mahanth Gowda

This paper introduces mmWave-Whisper, a system that demonstrates the feasibility of full-corpus automated speech recognition (ASR) on phone calls eavesdropped remotely using off-th…