most citedALR-GAN: Adaptive Layout Refinement for Text-to-Image Synthesis

30 citations · 35 across the 5 of their papers we have counts for

collaborators

5 papers

cs.RO2024

Physics-Aware Iterative Learning and Prediction of Saliency Map for Bimanual Grasp Planning

Shiyao Wang, Xiuping Liu, Charlie C. L. Wang +1

Learning the skill of human bimanual grasping can extend the capabilities of robotic systems when grasping large or heavy objects. However, it requires a much larger search space f…

cs.CV2024

Small Object Tracking in LiDAR Point Cloud: Learning the Target-awareness Prototype and Fine-grained Search Region

Shengjing Tian, Yinan Han, Xiuping Liu +1

Single Object Tracking in LiDAR point cloud is one of the most essential parts of environmental perception, in which small objects are inevitable in real-world scenarios and will b…

cs.CV20232 cited

AffordPose: A Large-scale Dataset of Hand-Object Interactions with Affordance-driven Hand Pose

Juntao Jian, Xiuping Liu, Manyi Li +2

How human interact with objects depends on the functional roles of the target objects, which introduces the problem of affordance-aware hand-object interaction. It requires a large…

cs.CV20233 cited

Fine-grained Text and Image Guided Point Cloud Completion with CLIP Model

Wei Song, Jun Zhou, Mingjie Wang +3

This paper focuses on the recently popular task of point cloud completion guided by multimodal information. Although existing methods have achieved excellent performance by fusing…

cs.CV202330 cited

ALR-GAN: Adaptive Layout Refinement for Text-to-Image Synthesis

Hongchen Tan, Baocai Yin, Kun Wei +2

We propose a novel Text-to-Image Generation Network, Adaptive Layout Refinement Generative Adversarial Network (ALR-GAN), to adaptively refine the layout of synthesized images with…