◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yi-Jie Huang

4 papers hereh-index 4128 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV3
  • cs.LG1

identity via Semantic Scholar / OpenAlex

most citedOpen-Set Image Tagging with Multi-Grained Text Supervision

3 citations · 3 across the 2 of their papers we have counts for

collaborators

4 papers

cs.CV2023★ 3 cited

Open-Set Image Tagging with Multi-Grained Text Supervision

Xinyu Huang, Yi-Jie Huang, Youcai Zhang +6

In this paper, we introduce the Recognize Anything Plus Model (RAM++), an open-set image tagging model effectively leveraging multi-grained text supervision. Previous approaches (e…

cs.CV2023

u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Jinjin Xu, Liwu Xu, Yuzhe Yang +5

Recent advancements in multi-modal large language models (MLLMs) have led to substantial improvements in visual understanding, primarily driven by sophisticated modality alignment…

cs.LG2023

Prototype Fission: Closing Set for Robust Open-set Semi-supervised Learning

Xuwei Tan, Yi-Jie Huang, Yaqian Li

Semi-supervised Learning (SSL) has been proven vulnerable to out-of-distribution (OOD) samples in realistic large-scale unsupervised datasets due to over-confident pseudo-labeling…

cs.CV2023

CLIP Brings Better Features to Visual Aesthetics Learners

Liwu Xu, Jinjin Xu, Yuzhe Yang +3

Image Aesthetics Assessment (IAA) is a challenging task due to its subjective nature and expensive manual annotations. Recent large-scale vision-language models, such as Contrastiv…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.