4 citations · 8 across the 6 of their papers we have counts for
4 papers · 1 filter
FiRe: Fixed-Noise Refinement for Visual Counterfactual Explanations
Yan Zeng, Changlu Guo, Oskar Kristoffersen +3
Visual counterfactual explanations aim to change classifier decisions through realistic and localized edits while preserving decision-irrelevant content. Existing DDPM-based method…
What Matters in Training a GPT4-Style Language Model with Multimodal Inputs?
Yan Zeng, Hanbo Zhang, Jiani Zheng +5
Recent advancements in Large Language Models (LLMs) such as GPT4 have displayed exceptional multi-modal capabilities in following open-ended instructions given images. However, the…
eTag: Class-Incremental Learning with Embedding Distillation and Task-Oriented Generation
Libo Huang, Yan Zeng, Chuanguang Yang +3
Class-Incremental Learning (CIL) aims to solve the neural networks' catastrophic forgetting problem, which refers to the fact that once the network updates on a new task, its perfo…
VLUE: A Multi-Task Benchmark for Evaluating Vision-Language Models
Wangchunshu Zhou, Yan Zeng, Shizhe Diao +1
Recent advances in vision-language pre-training (VLP) have demonstrated impressive performance in a range of vision-language (VL) tasks. However, there exist several challenges for…