3 papers
cs.CV2025
Image-POSER: Reflective RL for Multi-Expert Image Generation and Editing
Hossein Mohebbi, Mohammed Abdulrahman, Yanting Miao +2
Recent advances in text-to-image generation have produced strong single-shot models, yet no individual system reliably executes the long, compositional prompts typical of creative…
cs.AI2025
Constraints-Guided Diffusion Reasoner for Neuro-Symbolic Learning
Xuan Zhang, Zhijian Zhou, Weidi Xu +3
Enabling neural networks to learn complex logical constraints and fulfill symbolic reasoning is a critical challenge. Bridging this gap often requires guiding the neural network's…
cs.LG2025
A Minimalist Method for Fine-tuning Text-to-Image Diffusion Models
Yanting Miao, William Loh, Pacal Poupart +1
Recent work uses reinforcement learning (RL) to fine-tune text-to-image diffusion models, improving text-image alignment and sample quality. However, existing approaches introduce…