1 citations · 1 across the 1 of their papers we have counts for
4 papers
T2I-FineEval: Fine-Grained Compositional Metric for Text-to-Image Evaluation
Seyed Mohammad Hadi Hosseini, Amir Mohammad Izadi, Ali Abdollahi +2
Although recent text-to-image generative models have achieved impressive performance, they still often struggle with capturing the compositional complexities of prompts including a…
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation
Amir Mohammad Izadi, Seyed Mohammad Hadi Hosseini, Soroush Vafaie Tabar +3
Text-to-image generative models have made significant advancements in recent years; however, accurately capturing intricate details in textual prompts-such as entity missing, attri…
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
Armin Saghafian, Amirmohammad Izadi, Negin Hashemi Dijujin +1
Grounding the instruction in the environment is a key step in solving language-guided goal-reaching reinforcement learning problems. In automated reinforcement learning, a key conc…
ComAlign: Compositional Alignment in Vision-Language Models
Ali Abdollah, Amirmohammad Izadi, Armin Saghafian +5
Vision-language models (VLMs) like CLIP have showcased a remarkable ability to extract transferable features for downstream tasks. Nonetheless, the training process of these models…