activity
20222026
most citedOnline Foundation Model Selection in Robotics

1 citations · 3 across the 16 of their papers we have counts for

collaborators
Showing 2024Show all

6 papers · 1 filter

cs.LG2024

Dense Dynamics-Aware Reward Synthesis: Integrating Prior Experience with Demonstrations

Cevahir Koprulu, Po-han Li, Tianyu Qiu +5

Many continuous control problems can be formulated as sparse-reward reinforcement learning (RL) tasks. In principle, online RL methods can automatically explore the state space to…

cs.CV2024★ 1 cited

Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction

Po-han Li, Yunhao Yang, Mohammad Omama +2

Autonomous agents perceive and interpret their surroundings by integrating multimodal inputs, such as vision, audio, and LiDAR. These perceptual modalities support retrieval tasks,…

cs.LG2024

CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features

Po-han Li, Sandeep P. Chinchali, Ufuk Topcu

Multimodal encoders like CLIP excel in tasks such as zero-shot image classification and cross-modal retrieval. However, they require excessive training data. We propose canonical s…

cs.IR2024

Exploiting Distribution Constraints for Scalable and Efficient Image Retrieval

Mohammad Omama, Po-han Li, Sandeep P. Chinchali

Image retrieval is crucial in robotics and computer vision, with downstream applications in robot place recognition and vision-based product recommendations. Modern retrieval syste…

cs.RO2024

PEERNet: An End-to-End Profiling Tool for Real-Time Networked Robotic Systems

Aditya Narayanan, Pranav Kasibhatla, Minkyu Choi +3

Networked robotic systems balance compute, power, and latency constraints in applications such as self-driving vehicles, drone swarms, and teleoperated surgery. A core problem in t…

cs.RO2024★ 1 cited

Online Foundation Model Selection in Robotics

Po-han Li, Oyku Selin Toprak, Aditya Narayanan +2

Foundation models have recently expanded into robotics after excelling in computer vision and natural language processing. The models are accessible in two ways: open-source or pai…