clip 1continual learning 1cross-dataset consistency 1cross-modal retrieval 1efficiency 1evaluation framework 1multimodal representation 1robustness 1vision-language models 1weight interpolation 1
From the 2 of 3 linked papers with an AI index.
3 papers
cs.CV2026
AlphaWiSE: Adaptive Weight Interpolation for Continual Multimodal Representation Learning
Sarthak Jain, Qiran Hu, Zhen Zhu +1
The paper introduces AlphaWiSE, a post‑hoc weight‑space interpolation technique that combines two frozen checkpoints with learned scalar coefficients to improve continual learning…
cs.IR2026
Can Argus Judge Them All? Comparing VLMs Across Domains
Harsh Joshi, Gautam Siddharth Kashyap, Rafiq Ali +5
The paper introduces ARGUS-EVAL, a framework that assesses vision-language models on both capability and reliability across domains, and uses it to compare several VLMs on retrieva…
cs.CV2026
AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation
Eric Ji, Qiran Hu, Wufei Ma +4
Synthetic data generation has emerged as a powerful tool for improving data scalability in computer vision. Recent diffusion-based pipelines have demonstrated strong photorealism.…