4 papers
From Classification to Ranking: Enhancing LLM Reasoning Capabilities for MBTI Personality Detection
Yuan Cao, Feixiang Liu, Xinyue Wang +4
Personality detection aims to measure an individual's corresponding personality traits through their social media posts. The advancements in Large Language Models (LLMs) offer nove…
AlignVTOFF: Texture-Spatial Feature Alignment for High-Fidelity Virtual Try-Off
Yihan Zhu, Mengying Ge
Virtual Try-Off (VTOFF) is a challenging multimodal image generation task that aims to synthesize high-fidelity flat-lay garments under complex geometric deformation and rich high-…
Action-Dynamics Modeling and Cross-Temporal Interaction for Online Action Understanding
Xinyu Yang, Zheheng Jiang, Feixiang Zhou +5
Action understanding, encompassing action detection and anticipation, plays a crucial role in numerous practical applications. However, untrimmed videos are often characterized by…
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
Jiawen Lin, Shiran Bian, Yihang Zhu +4
3D Visual Grounding (3DVG) aims to localize objects in 3D scenes using natural language descriptions. Although supervised methods achieve higher accuracy in constrained settings, z…