2 papers
cs.CL2025
DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization
Zhenglin Zhou, Xiaobo Xia, Fan Ma +3
Text-to-3D generation automates 3D content creation from textual descriptions, which offers transformative potential across various fields. However, existing methods often struggle…
cs.CV2025
TV-Dialogue: Crafting Theme-Aware Video Dialogues with Immersive Interaction
Sai Wang, Fan Ma, Xinyi Li +2
Recent advancements in LLMs have accelerated the development of dialogue generation across text and images, yet video-based dialogue generation remains underexplored and presents u…