8 papers
A Brain-inspired Embodied Intelligence for Fluid and Fast Reflexive Robotics Control
Weiyu Guo, He Zhang, Pengteng Li +7
Recent advances in embodied intelligence have leveraged massive scaling of data and model parameters to master natural-language command following and multi-task control. In contras…
From Events to Clarity: The Event-Guided Diffusion Framework for Dehazing
Ling Wang, Yunfan Lu, Wenzong Ma +3
Clear imaging under hazy conditions is a critical task. Prior-based and neural methods have improved results. However, they operate on RGB frames, which suffer from limited dynamic…
Beyond Boundaries: Leveraging Vision Foundation Models for Source-Free Object Detection
Huizai Yao, Sicheng Zhao, Pengteng Li +6
Source-Free Object Detection (SFOD) aims to adapt a source-pretrained object detector to a target domain without access to source data. However, existing SFOD methods predominantly…
See&Trek: Training-Free Spatial Prompting for Multimodal Large Language Model
Pengteng Li, Pinhao Song, Wuyang Li +5
We introduce SEE&TREK, the first training-free prompting framework tailored to enhance the spatial understanding of Multimodal Large Language Models (MLLMS) under vision-only const…
From Events to Enhancement: A Survey on Event-Based Imaging Technologies
Yunfan Lu, Xiaogang Xu, Pengteng Li +4
Event cameras offering high dynamic range and low latency have emerged as disruptive technologies in imaging. Despite growing research on leveraging these benefits for different im…
Multimodal 3D Genome Pre-training
Minghao Yang, Pengteng Li, Yan Liang +6
Deep learning techniques have driven significant progress in various analytical tasks within 3D genomics in computational biology. However, a holistic understanding of 3D genomics…