9 citations · 10 across the 5 of their papers we have counts for
5 papers
A Multimodal Foundation Model to Enhance Generalizability and Data Efficiency for Pan-cancer Prognosis Prediction
Huajun Zhou, Fengtao Zhou, Jiabo Ma +6
Multimodal data provides heterogeneous information for a holistic understanding of the tumor microenvironment. However, existing AI models often struggle to harness the rich inform…
PerfCam: Digital Twinning for Production Lines Using 3D Gaussian Splatting and Vision Models
Michel Gokan Khan, Renan Guarese, Fabian Johnson +6
We introduce PerfCam, an open source Proof-of-Concept (PoC) digital twinning framework that combines camera and sensory data with 3D Gaussian Splatting and computer vision models f…
Exploration-Driven Generative Interactive Environments
Nedko Savov, Naser Kazemi, Mohammad Mahdi +3
Modern world models require costly and time-consuming collection of large video datasets with action demonstrations by people or by environment-specific agents. To simplify trainin…
Adaptive Visual Perception for Robotic Construction Process: A Multi-Robot Coordination Framework
Jia Xu, Manish Dixit, Xi Wang
Construction robots operate in unstructured construction sites, where effective visual perception is crucial for ensuring safe and seamless operations. However, construction robots…
SPARC-LoRa: A Scalable, Power-efficient, Affordable, Reliable, and Cloud Service-enabled LoRa Networking System for Agriculture Applications
Xi Wang, Bryan Hatasaka, Zhengyan Liu +10
With the rapid development of cloud and edge computing, Internet of Things (IoT) applications have been deployed in various aspects of human life. In this paper, we design and impl…