4 papers
OpenEAI-Platform: An Open-source Embodied Artificial Intelligence Hardware-Software Unified Platform
Jinyuan Zhang, Luoyi Fan, Leiyu Wang +4
Embodied AI in the real world requires both accurate hardware and robust vision-language-action (VLA) policies. We present OpenEAI-Platform, a fully open-source platform that integ…
Augur: Modeling Covariate Causal Associations in Time Series via Large Language Models
Zhiqing Cui, Binwu Wang, Qingxiang Liu +4
Large language models (LLM) have emerged as a promising avenue for time series forecasting, offering the potential to integrate multimodal data. However, existing LLM-based approac…
ViP-CLIP: Visual-Perception Prompting with Unified Alignment for Zero-Shot Anomaly Detection
Ziteng Yang, Jingzehua Xu, Yanshu Li +3
Zero-shot anomaly detection (ZSAD) aims to detect anomalies without any target domain training samples, relying solely on external auxiliary data. Existing CLIP-based methods attem…
LLaVA-RadZ: Can Multimodal Large Language Models Effectively Tackle Zero-shot Radiology Recognition?
Bangyan Li, Wenxuan Huang, Zhenkun Gao +8
Recently, Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in visual understanding and reasoning across various vision-language tasks. However, w…