2 papers
cs.CV2024
VideoSAM: Open-World Video Segmentation
Pinxue Guo, Zixu Zhao, Jianxiong Gao +5
Video segmentation is essential for advancing robotics and autonomous driving, particularly in open-world settings where continuous perception and object association across video f…
cs.CV2024
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
Jiaxin Cheng, Zixu Zhao, Tong He +3
Recent advancements in generative models have significantly enhanced their capacity for image generation, enabling a wide range of applications such as image editing, completion an…