12 papers
Foundation-Assisted Active Learning for Object Detection Annotation
Jinchang Zhang, Arnold Zumbrun, Jing Lin +1
The annotation cost for remote sensing object detection is high, while existing active learning methods still face several challenges in object detection scenarios, including the c…
EmbodiedUS-FS: Fast Slow Intelligence for Ultrasound Robotics
Fangzhuo Zhang, Xinyu Wang, Xiao Yang +1
Robotic ultrasound scanning in real clinical environments requires both high-level clinical workflow reasoning and low-level closed-loop execution. Physicians natural-language inst…
Interpretable Traffic Responsibility from Dashcam Video via Legal Multi Agent Reasoning
Jingchun Yang, Jinchang Zhang
The widespread adoption of dashcams has made video evidence in traffic accidents increasingly abundant, yet transforming "what happened in the video" into "who is responsible under…
Fractal Autoregressive Depth Estimation with Continuous Token Diffusion
Jinchang Zhang, Xinrou Kang, Guoyu Lu
Monocular depth estimation can benefit from autoregressive (AR) generation, but direct AR modeling is hindered by the modality gap between RGB and depth, inefficient pixel-wise gen…
Automated Genomic Interpretation via Concept Bottleneck Models for Medical Robotics
Zijun Li, Jinchang Zhang, Ming Zhang +1
We propose an automated genomic interpretation module that transforms raw DNA sequences into actionable, interpretable decisions suitable for integration into medical automation an…
Adaptive Event Stream Slicing for Open-Vocabulary Event-Based Object Detection via Vision-Language Knowledge Distillation
Jinchang Zhang, Zijun Li, Jiakai Lin +1
Event cameras offer advantages in object detection tasks due to high-speed response, low latency, and robustness to motion blur. However, event cameras lack texture and color infor…