high-resolution image synthesis 2masked diffusion 2controllable generation 1diffusion models 1discrete token generation 1image editing 1multimodal diffusion 1multimodal understanding 1object grounding 1parallel decoding 1text-to-image 1token editing 1
From the 4 of 20 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
MobileWorldBench: Towards Semantic World Modeling For Mobile Agents
Shufan Li, Konstantinos Kallidromitis, Akash Gokul +3
World models have shown great utility in improving the task performance of embodied agents. While prior work largely focuses on pixel-space world models, these approaches face prac…
cs.AI2025
MedMax: Mixed-Modal Instruction Tuning for Training Biomedical Assistants
Hritik Bansal, Daniel Israel, Siyan Zhao +3
Recent advancements in mixed-modal generative have opened new avenues for developing unified biomedical assistants capable of analyzing biomedical images, answering complex questio…