activity
20182026
most citedNeoBabel: A Multilingual Open Tower for Visual Generation

1 citations · 3 across the 11 of their papers we have counts for

collaborators

11 papers

cs.CV2026

FineHOI: Part-Aware Dense Representations for Zero-Shot Human-Object Interaction Detection

Francesco Tonini, Lorenzo Vaquero, Mohammad Mahdi Derakhshani +3

Human-Object Interaction (HOI) detection aims to localize humans and objects in images and classify their interactions. Zero-shot HOI focuses on recognizing interactions that are n…

cs.CV2026

GAP3D: Generative Alignment of VLM Latents to Patch-Level Embeddings for 3D Generation

Polytimi Anna Gkotsi, Andrii Zadaianchuk, Mohammad Mahdi Derakhshani

Recent approaches integrating vision-language models (VLMs) as prompt encoders for generative model conditioning typically rely on expensive end-to-end training or map features to…

cs.LG2026

Private PoEtry: Private In-Context Learning via Product of Experts

Rob Romijnders, Mohammad Mahdi Derakhshani, Jonathan Petit +3

In-context learning (ICL) enables Large Language Models (LLMs) to adapt to new tasks with only a small set of examples at inference time, thereby avoiding task-specific fine-tuning…

cs.CV2025

Purrception: Variational Flow Matching for Vector-Quantized Image Generation

Răzvan-Andrei Matişan, Vincent Tao Hu, Grigory Bartosh +6

We introduce Purrception, a variational flow matching approach for vector-quantized image generation that provides explicit categorical supervision while maintaining continuous tra…

cs.CL2025

NeoBabel: A Multilingual Open Tower for Visual Generation

Mohammad Mahdi Derakhshani, Dheeraj Varghese, Marzieh Fadaee +1

Text-to-image generation advancements have been predominantly English-centric, creating barriers for non-English speakers and perpetuating digital inequities. While existing system…

cs.CV2025

Continual Hyperbolic Learning of Instances and Classes

Melika Ayoughi, Mina Ghadimi Atigh, Mohammad Mahdi Derakhshani +3

Continual learning has traditionally focused on classifying either instances or classes, but real-world applications, such as robotics and self-driving cars, require models to hand…