6 papers
MapReduce LoRA: Advancing the Pareto Front in Multi-Preference Optimization for Generative Models
Chieh-Yun Chen, Zhonghao Wang, Qi Chen +10
Reinforcement learning from human feedback (RLHF) with reward models has advanced alignment of generative models to human aesthetic and perceptual preferences. However, jointly opt…
Assessing Image Quality Issues for Real-World Problems
Tai-Yin Chiu, Yinan Zhao, Danna Gurari
We introduce a new large-scale dataset that links the assessment of image quality issues to two practical vision tasks: image captioning and visual question answering. First, we id…
Captioning Images Taken by People Who Are Blind
Danna Gurari, Yinan Zhao, Meng Zhang +1
While an important problem in the vision community is to design algorithms that can automatically caption images, few publicly-available datasets for algorithm development directly…
Unconstrained Foreground Object Search
Yinan Zhao, Brian Price, Scott Cohen +1
Many people search for foreground objects to use when editing images. While existing methods can retrieve candidates to aid in this, they are constrained to returning objects that…
Predicting How to Distribute Work Between Algorithms and Humans to Segment an Image Batch
Danna Gurari, Yinan Zhao, Suyog Dutt Jain +2
Foreground object segmentation is a critical step for many image analysis tasks. While automated methods can produce high-quality results, their failures disappoint users in need o…
Guided Image Inpainting: Replacing an Image Region by Pulling Content from Another Image
Yinan Zhao, Brian Price, Scott Cohen +1
Deep generative models have shown success in automatically synthesizing missing image regions using surrounding context. However, users cannot directly decide what content to synth…