interest modeling 1knowledge distillation 1latent reasoning 1lightweight models 1multimodal emotion recognition 1multimodal large language models 1optimal transport loss 1recursive reasoning 1sequential recommendation 1supervised training 1
From the 2 of 19 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Focal-RegionFace: Generating Fine-Grained Multi-attribute Descriptions for Arbitrarily Selected Face Focal Regions
Kaiwen Zheng, Junchen Fu, Songpei Xu +4
In this paper, we introduce an underexplored problem in facial analysis: generating and recognizing multi-attribute natural language descriptions, containing facial action units (A…
cs.CV2025
Video-Bench: Human-Aligned Video Generation Benchmark
Hui Han, Siyuan Li, Jiaqi Chen +10
Video generation assessment is essential for ensuring that generative models produce visually realistic, high-quality videos while aligning with human expectations. Current video g…
cs.CV2025
Multimodal Representation Learning Techniques for Comprehensive Facial State Analysis
Kaiwen Zheng, Xuri Ge, Junchen Fu +2
Multimodal foundation models have significantly improved feature representation by integrating information from multiple modalities, making them highly suitable for a broader set o…