3 papers
cs.CV2025
Motus: A Unified Latent Action World Model
Hongzhe Bi, Hengkai Tan, Shenghao Xie +13
While a general embodied agent must function as a unified system, current methods are built on isolated models for understanding, world modeling, and control. This fragmentation pr…
cs.CV2025
Synthesizing Reality: Leveraging the Generative AI-Powered Platform Midjourney for Construction Worker Detection
Hongyang Zhao, Tianyu Liang, Sina Davari +1
While recent advancements in deep neural networks (DNNs) have substantially enhanced visual AI's capabilities, the challenge of inadequate data diversity and volume remains, partic…
cs.CL2025
Mixed-Precision Graph Neural Quantization for Low Bit Large Language Models
Wanlong Liu, Yichen Xiao, Dingyi Zeng +3
Post-Training Quantization (PTQ) is pivotal for deploying large language models (LLMs) within resource-limited settings by significantly reducing resource demands. However, existin…