3 papers
cs.CV2025
VideoPerceiver: Enhancing Fine-Grained Temporal Perception in Video Multimodal Large Language Models
Fufangchen Zhao, Liao Zhang, Daiqi Shi +5
We propose VideoPerceiver, a novel video multimodal large language model (VMLLM) that enhances fine-grained perception in video understanding, addressing VMLLMs' limited ability to…
cs.CL2025
SitLLM: Large Language Models for Sitting Posture Health Understanding via Pressure Sensor Data
Jian Gao, Fufangchen Zhao, Yiyang Zhang +1
Poor sitting posture is a critical yet often overlooked factor contributing to long-term musculoskeletal disorders and physiological dysfunctions. Existing sitting posture monitori…
cs.CL2024
reCSE: Portable Reshaping Features for Sentence Embedding in Self-supervised Contrastive Learning
Fufangchen Zhao, Jian Gao, Danfeng Yan
We propose reCSE, a self supervised contrastive learning sentence representation framework based on feature reshaping. This framework is different from the current advanced models…