3 papers
cs.CV2025
AdaViP: Aligning Multi-modal LLMs via Adaptive Vision-enhanced Preference Optimization
Jinda Lu, Jinghan Li, Yuan Gao +4
Preference alignment through Direct Preference Optimization (DPO) has demonstrated significant effectiveness in aligning multimodal large language models (MLLMs) with human prefere…
cs.LG2024
Adaptive Self-supervised Robust Clustering for Unstructured Data with Unknown Cluster Number
Chen-Lu Ding, Jiancan Wu, Wei Lin +3
We introduce a novel self-supervised deep clustering approach tailored for unstructured data without requiring prior knowledge of the number of clusters, termed Adaptive Self-super…
cs.CV2023
Text-to-Image Generation for Abstract Concepts
Jiayi Liao, Xu Chen, Qiang Fu +5
Recent years have witnessed the substantial progress of large-scale models across various domains, such as natural language processing and computer vision, facilitating the express…