papers

Publications (5)

cs.CV2026

UMSS: Towards Unsupervised Multi-modal Semantic Segmentation

Haitian Zhang, Thai Duy Nguyen, Xiangyuan Wang +2

The paper introduces UniM2, an unsupervised framework for multimodal semantic segmentation that learns a shared latent space across sensors using cross‑modal correspondence and a h…

#unsupervised semantic segmentation#multimodal fusion#cross-modal learning#depth perception
cs.RO2026

Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains

Xuanyu Chen, Mohan Liu, Dengchen Mei +6

Model-free reinforcement learning has enabled impressive humanoid locomotion; however, control on steep slopes remains largely unexplored. Unlike flat or discrete terrains, sloped…

cs.RO2026

VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation

Mohan Liu, Zhihao Gu, Xuanyu Chen +5

VistaVLA is a two-stage framework that builds a geometry- and semantic-aware 3D cognitive map using Gaussian primitives and compresses it into compact tokens for vision‑language‑ac…

#vision-language-action#3d representation#gaussian primitives#semantic grounding
cond-mat.mtrl-sci2021

OPTIMADE, an API for exchanging materials data

Casper W. Andersen, Rickard Armiento, Evgeny Blokhin +53

The Open Databases Integration for Materials Design (OPTIMADE) consortium has designed a universal application programming interface (API) to make materials databases accessible an…

cs.CV2026

Rad-VLSM: A Cross-Modal Framework with Semantics-Assisted Prompting for Medical Segmentation and Diagnosis

Fengyi Zhang, Xujie Zeng, Mohan Liu +2

Medical image segmentation is more clinically valuable when it supports diagnosis rather than merely producing lesion masks. However, diagnostically relevant lesion cues are often…