curriculum learning 1dataset construction 1evaluation metrics 1text-to-image generation 1vision-language models 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
DynEval: Holistic Evaluations of T2I Generative Models in the Wild
Shyam Marjit, Dheeraj Baiju, Anuj Shikarkhane +3
The paper introduces DynEval, a dynamic evaluation framework that jointly assesses text-to-image alignment and image quality for T2I models, using large synthetic datasets and a di…
cs.CV2026
Prompt Estimation from Prototypes for Federated Prompt Tuning of Vision Transformers
M Yashwanth, Sharannya Ghosh, Aditay Tripathi +1
Visual Prompt Tuning (VPT) of pre-trained Vision Transformers (ViTs) has proven highly effective as a parameter-efficient fine-tuning technique for adapting large models to downstr…
cs.CV2025
O3SLM: Open Weight, Open Data, and Open Vocabulary Sketch-Language Model
Rishi Gupta, Mukilan Karuppasamy, Shyam Marjit +2
While Large Vision Language Models (LVLMs) are increasingly deployed in real-world applications, their ability to interpret abstract visual inputs remains limited. Specifically, th…