1 citations · 3 across the 6 of their papers we have counts for
6 papers
SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions
Xiaoyu Liu, Yuxiang Wei, Ming Liu +4
Human visual imagination usually begins with analogies or rough sketches. For example, given an image with a girl playing guitar before a building, one may analogously imagine how…
Distilling Semantic Priors from SAM to Efficient Image Restoration Models
Quan Zhang, Xiaoyu Liu, Wei Li +6
In image restoration (IR), leveraging semantic priors from segmentation models has been a common approach to improve performance. The recent segment anything model (SAM) has emerge…
Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences
Xiyao Wang, Yuhang Zhou, Xiaoyu Liu +9
Multimodal Large Language Models (MLLMs) have demonstrated proficiency in handling a variety of visual-language tasks. However, current MLLM benchmarks are predominantly designed t…
Error Correlations in Photonic Qudit-Mediated Entanglement Generation
Xiaoyu Liu, Niv Bharos, Liubov Markovich +1
Generating entanglement between distributed network nodes is a prerequisite for the quantum internet. Entanglement distribution protocols based on high-dimensional photonic qudits…
C-Disentanglement: Discovering Causally-Independent Generative Factors under an Inductive Bias of Confounder
Xiaoyu Liu, Jiaxin Yuan, Bang An +3
Representation learning assumes that real-world data is generated by a few semantically meaningful generative factors (i.e., sources of variation) and aims to discover them in the…
Beyond Image Borders: Learning Feature Extrapolation for Unbounded Image Composition
Xiaoyu Liu, Ming Liu, Junyi Li +4
For improving image composition and aesthetic quality, most existing methods modulate the captured images by striking out redundant content near the image borders. However, such im…