2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Digi-Q: Learning Q-Value Functions for Training Device-Control Agents
Hao Bai, Yifei Zhou, Li Erran Li +2
While a number of existing approaches for building foundation model agents rely on prompting or fine-tuning with human demonstrations, it is not sufficient in dynamic environments…
cs.RO2024★ 2 cited
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
Grace Tang, Swetha Rajkumar, Yifei Zhou +3
Building generalist robotic systems involves effectively endowing robots with the capabilities to handle novel objects in an open-world setting. Inspired by the advances of large p…