papers

Publications (11)

q-bio.GN2026

GenoBERT: A Language Model for Accurate Genotype Imputation

Lei Huang, Chuan Qiu, Kuan-Jui Su +13

Genotype imputation enables dense variant coverage for genome-wide association and risk-prediction studies, yet conventional reference-panel methods remain limited by ancestry bias…

cs.CV2024

CSS: Overcoming Pose and Scene Challenges in Crowd-Sourced 3D Gaussian Splatting

Runze Chen, Mingyu Xiao, Haiyong Luo +5

We introduce Crowd-Sourced Splatting (CSS), a novel 3D Gaussian Splatting (3DGS) pipeline designed to overcome the challenges of pose-free scene reconstruction using crowd-sourced…

cs.LG2024

A Minimalist Prompt for Zero-Shot Policy Learning

Meng Song, Xuezhi Wang, Tanay Biradar +2

Transformer-based methods have exhibited significant generalization ability when prompted with target-domain demonstrations or example solutions during inference. Although demonstr…

eess.SY2026

Exergy Battery Modeling and P2P Trading Based Optimal Operation of Virtual Energy Station

Meng Song, Xinyi Jing, Jianyong Ding +4

Virtual energy stations (VESs) work as retailers to provide electricity and natural gas sale services for integrated energy systems (IESs), and guide IESs energy consumption behavi…

cs.RO2022

Learning to Rearrange with Physics-Inspired Risk Awareness

Meng Song, Yuhan Liu, Zhengqin Li +1

Real-world applications require a robot operating in the physical world with awareness of potential risks besides accomplishing the task. A large part of risky behaviors arises fro…

cs.LG2025

Good Actions Succeed, Bad Actions Generalize: A Case Study on Why RL Generalizes Better

Meng Song

Supervised learning (SL) and reinforcement learning (RL) are both widely used to train general-purpose agents for complex tasks, yet their generalization capabilities and underlyin…

cs.CV2026

SSR: Pushing the Limit of Spatial Intelligence with Structured Scene Reasoning

Yi Zhang, Youya Xia, Yong Wang +7

While Multimodal Large Language Models (MLLMs) excel in semantic tasks, they frequently lack the "spatial sense" essential for sophisticated geometric reasoning. Current models typ…

cs.CV2021

OpenRooms: An End-to-End Open Framework for Photorealistic Indoor Scene Datasets

Zhengqin Li, Ting-Wei Yu, Shen Sang +14

We propose a novel framework for creating large-scale photorealistic datasets of indoor scenes, with ground truth geometry, material, lighting and semantics. Our goal is to make th…

cs.LG2024

Probabilistic World Modeling with Asymmetric Distance Measure

Meng Song

Representation learning is a fundamental task in machine learning, aiming at uncovering structures from data to facilitate subsequent tasks. However, what is a good representation…

cs.CL2022

RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning

Mingkai Deng, Jianyu Wang, Cheng-Ping Hsieh +6

Prompting has shown impressive success in enabling large pretrained language models (LMs) to perform diverse NLP tasks, especially when only few downstream data are available. Auto…

cs.RO2019

S4G: Amodal Single-view Single-Shot SE(3) Grasp Detection in Cluttered Scenes

Yuzhe Qin, Rui Chen, Hao Zhu +3

Grasping is among the most fundamental and long-lasting problems in robotics study. This paper studies the problem of 6-DoF(degree of freedom) grasping by a parallel gripper in a c…