43 citations · 43 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Towards Unified Facial Action Unit Recognition Framework by Large Language Models
Guohong Hu, Xing Lan, Hanyu Jiang +2
Facial Action Units (AUs) are of great significance in the realm of affective computing. In this paper, we propose AU-LLaVA, the first unified AU recognition framework based on the…
cs.CV2024
MVLLaVA: An Intelligent Agent for Unified and Flexible Novel View Synthesis
Hanyu Jiang, Jian Xue, Xing Lan +2
This paper introduces MVLLaVA, an intelligent agent designed for novel view synthesis tasks. MVLLaVA integrates multiple multi-view diffusion models with a large multimodal model,…
cs.CV2023★ 43 cited
FoodSAM: Any Food Segmentation
Xing Lan, Jiayi Lyu, Hanyu Jiang +4
In this paper, we explore the zero-shot capability of the Segment Anything Model (SAM) for food image segmentation. To address the lack of class-specific information in SAM-generat…