4 papers
A Chain-of-Thought Subspace Meta-Learning for Few-shot Image Captioning with Large Vision and Language Models
Hao Huang, Shuaihang Yuan, Yu Hao +2
A large-scale vision and language model that has been pretrained on massive data encodes visual and linguistic prior, which makes it easier to generate images and language that are…
Integrating Retrospective Framework in Multi-Robot Collaboration
Jiazhao Liang, Hao Huang, Yu Hao +4
Recent advancements in Large Language Models (LLMs) have demonstrated substantial capabilities in enhancing communication and coordination in multi-robot systems. However, existing…
Curvature Diversity-Driven Deformation and Domain Alignment for Point Cloud
Mengxi Wu, Hao Huang, Yi Fang +1
Unsupervised Domain Adaptation (UDA) is crucial for reducing the need for extensive manual data annotation when training deep networks on point cloud data. A significant challenge…
MultiTalk: Introspective and Extrospective Dialogue for Human-Environment-LLM Alignment
Venkata Naren Devarakonda, Ali Umut Kaypak, Shuaihang Yuan +3
LLMs have shown promising results in task planning due to their strong natural language understanding and reasoning capabilities. However, issues such as hallucinations, ambiguitie…