Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
Ben Vardi, Oron Nir, Ariel Shamir
Vision-Language Models (VLMs) demonstrate remarkable capabilities in visual understanding and reasoning, such as in Visual Question Answering (VQA), where the model is asked a ques…
cs.CV2024
Conditional Balance: Improving Multi-Conditioning Trade-Offs in Image Generation
Nadav Z. Cohen, Oron Nir, Ariel Shamir
Balancing content fidelity and artistic style is a pivotal challenge in image generation. While traditional style transfer methods and modern Denoising Diffusion Probabilistic Mode…
cs.CV2022
CAST: Character labeling in Animation using Self-supervision by Tracking
Oron Nir, Gal Rapoport, Ariel Shamir
Cartoons and animation domain videos have very different characteristics compared to real-life images and videos. In addition, this domain carries a large variability in styles. Cu…