hierarchical models 1multimodal financial data 1synthetic data generation 1tabular data 1top-down bottom-up 1
From the 1 of 6 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
A Generalized Label Shift Perspective for Cross-Domain Gaze Estimation
Hao-Ran Yang, Xiaohui Chen, Chuan-Xian Ren
Aiming to generalize the well-trained gaze estimation model to new target domains, Cross-domain Gaze Estimation (CDGE) is developed for real-world application scenarios. Existing C…
cs.CV2025
Multi-Modal Foundation Models for Computational Pathology: A Survey
Dong Li, Guihong Wan, Xintao Wu +7
Foundation models have emerged as a powerful paradigm in computational pathology (CPath), enabling scalable and generalizable analysis of histopathological images. While early deve…
cs.CV2024
CompCap: Improving Multimodal Large Language Models with Composite Captions
Xiaohui Chen, Satya Narayan Shukla, Mahmoud Azab +8
How well can Multimodal Large Language Models (MLLMs) understand composite images? Composite images (CIs) are synthetic visuals created by merging multiple visual elements, such as…