From the 1 of 6 linked papers with an AI index.
4 papers · 1 filter
Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling
Bo-An Chang, Yu-Chih Chen
As Text-to-Image (T2I) systems rapidly advance, evaluating the cultural authenticity of synthesized content has become increasingly important for fair and trustworthy generative AI…
Bring Music The Horizon: Music-Driven 360 Video Generation
Kai Hsu Tsai, Yong Wei Fu, Hung I Yang +1
The paper presents an emotion-aware pipeline that converts a song into an immersive 360° video by estimating its emotional trajectory, generating emotion-guided keyframes, and synt…
InfVSR: Toward Consistency-Driven Streaming Generative Video Super-Resolution
Ziqing Zhang, Kai Liu, Zheng Chen +5
Real-world videos often extend over thousands of frames. Existing generative video super-resolution (VSR) approaches, however, face two persistent challenges when processing long s…
Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressive Diffusion
Hau-Shiang Shiu, Chin-Yang Lin, Zhixiang Wang +4
Diffusion-based video super-resolution (VSR) methods deliver strong perceptual quality but are often unsuitable for latency-sensitive scenarios due to reliance on future frames and…