activity
20172025
most citedGLAD: Global-Local-Alignment Descriptor for Pedestrian Retrieval

425 citations · 597 across the 34 of their papers we have counts for

collaborators
Showing 2023 · cs.CVShow all

8 papers · 2 filters

cs.CV2023

Boosting Segment Anything Model Towards Open-Vocabulary Learning

Xumeng Han, Longhui Wei, Xuehui Yu +6

The recent Segment Anything Model (SAM) has emerged as a new paradigmatic vision foundation model, showcasing potent zero-shot generalization and flexible prompting. Despite SAM fi…

cs.CV2023★ 1 cited

HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data

Qifan Yu, Juncheng Li, Longhui Wei +6

Multi-modal Large Language Models (MLLMs) tuned on machine-generated instruction-following data have demonstrated remarkable performance in various multi-modal understanding and ge…

cs.CV2023

Degeneration-Tuning: Using Scrambled Grid shield Unwanted Concepts from Stable Diffusion

Zixuan Ni, Longhui Wei, Jiacheng Li +3

Owing to the unrestricted nature of the content in the training data, large text-to-image diffusion models, such as Stable Diffusion (SD), are capable of generating images with pot…

cs.CV2023★ 4 cited

SDDM: Score-Decomposed Diffusion Models on Manifolds for Unpaired Image-to-Image Translation

Shikun Sun, Longhui Wei, Junliang Xing +2

Recent score-based diffusion models (SBDMs) show promising results in unpaired image-to-image translation (I2I). However, existing methods, either energy-based or statistically-bas…

cs.CV2023★ 5 cited

Towards AGI in Computer Vision: Lessons Learned from GPT and Large Language Models

Lingxi Xie, Longhui Wei, Xiaopeng Zhang +4

The AI community has been pursuing algorithms known as artificial general intelligence (AGI) that apply to any kind of real-world problem. Recently, chat systems powered by large l…

cs.CV2023★ 2 cited

Learning Transferable Pedestrian Representation from Multimodal Information Supervision

Liping Bao, Longhui Wei, Xiaoyu Qiu +3

Recent researches on unsupervised person re-identification~(reID) have demonstrated that pre-training on unlabeled person images achieves superior performance on downstream reID ta…