1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
Scaffolding Coordinates to Promote Vision-Language Coordination in Large Multi-Modal Models
Xuanyu Lei, Zonghan Yang, Xinrui Chen +2
State-of-the-art Large Multi-Modal Models (LMMs) have demonstrated exceptional capabilities in vision-language tasks. Despite their advanced functionalities, the performances of LM…
cs.AI2024
Towards Unified Alignment Between Agents, Humans, and Environment
Zonghan Yang, An Liu, Zijun Liu +11
The rapid progress of foundation models has led to the prosperity of autonomous agents, which leverage the universal capabilities of foundation models to conduct reasoning, decisio…
cs.CV2023★ 1 cited
A XGBoost Algorithm-based Fatigue Recognition Model Using Face Detection
Xinrui Chen, Bingquan Zhang
As fatigue is normally revealed in the eyes and mouth of a person's face, this paper tried to construct a XGBoost Algorithm-Based fatigue recognition model using the two indicators…