most citedHierarchical Attention Network for Action Recognition in Videos

78 citations · 78 across the 1 of their papers we have counts for

collaborators

10 papers

cs.CV2024

YouTube SFV+HDR Quality Dataset

Yilin Wang, Joong Gon Yim, Neil Birkbeck +1

The popularity of Short form videos (SFV) has grown dramatically in the past few years, and has become a phenomenal video category with billions of viewers. Meanwhile, High Dynamic…

cs.CV2024

OmniControlNet: Dual-stage Integration for Conditional Image Generation

Yilin Wang, Haiyang Xu, Xiang Zhang +4

We provide a two-way integration for the widely adopted ControlNet by integrating external condition generation algorithms into a single dense prediction method and incorporating i…

cs.AI20246 cited

KC-GenRe: A Knowledge-constrained Generative Re-ranking Method Based on Large Language Models for Knowledge Graph Completion

Yilin Wang, Minghao Hu, Zhen Huang +3

The goal of knowledge graph completion (KGC) is to predict missing facts among entities. Previous methods for KGC re-ranking are mostly built on non-generative language models to o…

cs.CV20231 cited

Dolfin: Diffusion Layout Transformers without Autoencoder

Yilin Wang, Zeyuan Chen, Liangjun Zhong +3

In this paper, we introduce a novel generative model, Diffusion Layout Transformers without Autoencoder (Dolfin), which significantly improves the modeling capability with reduced…

cs.CV20237 cited

Photoswap: Personalized Subject Swapping in Images

Jing Gu, Yilin Wang, Nanxuan Zhao +8

In an era where images and visual content dominate our digital landscape, the ability to manipulate and personalize these images has become a necessity. Envision seamlessly substit…

cs.CV2023

MRET: Multi-resolution Transformer for Video Quality Assessment

Junjie Ke, Tianhao Zhang, Yilin Wang +2

No-reference video quality assessment (NR-VQA) for user generated content (UGC) is crucial for understanding and improving visual experience. Unlike video recognition tasks, VQA ta…