2 papers
cs.CV2025
VLDrive: Vision-Augmented Lightweight MLLMs for Efficient Language-grounded Autonomous Driving
Ruifei Zhang, Wei Zhang, Xiao Tan +4
Recent advancements in language-grounded autonomous driving have been significantly promoted by the sophisticated cognition and reasoning capabilities of large language models (LLM…
cs.CV2025
LGM-Pose: A Lightweight Global Modeling Network for Real-time Human Pose Estimation
Biao Guo, Cong Zhou, Fangmin Guo +3
Most of the current top-down multi-person pose estimation lightweight methods are based on multi-branch parallel pure CNN network architecture, which often struggle to capture the…