10 citations · 10 across the 3 of their papers we have counts for
3 papers
cs.CV2025
TACO: Think-Answer Consistency for Optimized Long-Chain Reasoning and Efficient Data Learning via Reinforcement Learning in LVLMs
Zhehan Kan, Yanlin Liu, Kun Yin +8
DeepSeek R1 has significantly advanced complex reasoning for large language models (LLMs). While recent methods have attempted to replicate R1's reasoning capabilities in multimoda…
cs.CV2024
TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling
Weiran Chen, Xin Li, Jiaqi Su +4
As a cross-modal task, visual storytelling aims to generate a story for an ordered image sequence automatically. Different from the image captioning task, visual storytelling requi…
cs.CV2023★ 10 cited
MorphGANFormer: Transformer-based Face Morphing and De-Morphing
Na Zhang, Xudong Liu, Xin Li +1
Semantic face image manipulation has received increasing attention in recent years. StyleGAN-based approaches to face morphing are among the leading techniques; however, they often…