2 papers
cs.CL2024
Social Debiasing for Fair Multi-modal LLMs
Harry Cheng, Yangyang Guo, Qingpei Guo +4
Multi-modal Large Language Models (MLLMs) have dramatically advanced the research field and delivered powerful vision-language understanding capabilities. However, these models oft…
cs.CV2023
Sample Less, Learn More: Efficient Action Recognition via Frame Feature Restoration
Harry Cheng, Yangyang Guo, Liqiang Nie +2
Training an effective video action recognition model poses significant computational challenges, particularly under limited resource budgets. Current methods primarily aim to eithe…