Publications (20)
Distilling Neuron Spike with High Temperature in Reinforcement Learning Agents
Ling Zhang, Jian Cao, Yuan Zhang +2
Spiking neural network (SNN), compared with depth neural network (DNN), has faster processing speed, lower energy consumption and more biological interpretability, which is expecte…
Efficient and Exact Multimarginal Optimal Transport with Pairwise Costs
Bohan Zhou, Matthew Parno
In this paper, we address the numerical solution to the multimarginal optimal transport (MMOT) with pairwise costs. MMOT, as a natural extension from the classical two-marginal opt…
Connections between Bressan's Mixing Conjecture, the Branched Optimal Transport and Combinatorial Optimization
Bohan Zhou
We investigate the 1D version of the notable Bressan's mixing conjecture, and introduce various formulation in the classical optimal transport setting, the branched optimal transpo…
DARA: Few-shot Budget Allocation in Online Advertising via In-Context Decision Making with RL-Finetuned LLMs
Mingxuan Song, Yusen Huo, Bohan Zhou +5
Optimizing the advertiser's cumulative value of winning impressions under budget constraints poses a complex challenge in online advertising, under the paradigm of AI-Generated Bid…
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
Bohan Zhou, Yi Zhan, Zhongbin Zhang +1
Egocentric hand-object motion generation is crucial for immersive AR/VR and robotic imitation but remains challenging due to unstable viewpoints, self-occlusions, perspective disto…
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
Haoqi Yuan, Yu Bai, Yuhui Fu +6
Building autonomous robotic agents capable of achieving human-level performance in real-world embodied tasks is an ultimate goal in humanoid robot research. Recent advances have ma…
Pre-trained Visual Dynamics Representations for Efficient Policy Learning
Hao Luo, Bohan Zhou, Zongqing Lu
Pre-training for Reinforcement Learning (RL) with purely video data is a valuable yet challenging problem. Although in-the-wild videos are readily available and inhere a vast amoun…
Learning from Visual Observation via Offline Pretrained State-to-Go Transformer
Bohan Zhou, Ke Li, Jiechuan Jiang +1
Learning from visual observation (LfVO), aiming at recovering policies from only visual observation data, is promising yet a challenging problem. Existing LfVO approaches either on…
The existence of minimizers for an isoperimetric problem with Wasserstein penalty term in unbounded domains
Qinglan Xia, Bohan Zhou
In this article, we consider the (double) minimization problem $$\min\left\{P(E;Ω)+λW_p(E,F):~E\subseteqΩ,~F\subseteq \mathbb{R}^d,~\lvert E\cap F\rvert=0,~ \lvert E\rvert=\lver…
Cloud-top infrared observations reveal the four-dimensional precipitation structure
Tianchi Xu, Ziqiang Ma, Andrea Marinoni +11
Accurate four-dimensional (4D) precipitation information is essential for understanding the Earth's energy and water cycles, yet remains observationally unresolved at global scales…
The Signed Wasserstein Barycenter Problem
Matt Jacobs, Bohan Zhou
Barycenter problems encode important geometric information about a metric space. While these problems are typically studied with positive weight coefficients associated to each dis…
Micromachining & FBG fabrication using point by point technique utilizing femto-second laser
Abu Farzan Mitul, Bohan Zhou, Huiyu Zhao +1
Fiber bragg gratings (FBG) has wide variety of applications in sensor and laser devices. In this work, we have fabricated FBG using point by point (PbP) technique utilizing fs lase…
Cradle: Empowering Foundation Agents Towards General Computer Control
Weihao Tan, Wentao Zhang, Xinrun Xu +25
Despite the success in specific scenarios, existing foundation agents still struggle to generalize across various virtual scenarios, mainly due to the dramatically different encaps…
Cross-Embodiment Dexterous Grasping with Reinforcement Learning
Haoqi Yuan, Bohan Zhou, Yuhui Fu +1
Dexterous hands exhibit significant potential for complex real-world grasping tasks. While recent studies have primarily focused on learning policies for specific robotic hands, th…
Learning Diverse Bimanual Dexterous Manipulation Skills from Human Demonstrations
Bohan Zhou, Haoqi Yuan, Yuhui Fu +1
Bimanual dexterous manipulation is a critical yet underexplored area in robotics. Its high-dimensional action space and inherent task complexity present significant challenges for…
Accelerated Markov Chain Monte Carlo Algorithms on Discrete States
Bohan Zhou, Shu Liu, Xinzhe Zuo +1
We propose a class of discrete state sampling algorithms based on Nesterov's accelerated gradient method, which extends the classical Metropolis-Hastings (MH) algorithm. The evolut…
Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence Analysis
Kaheon Kim, Bohan Zhou, Changbo Zhu +1
This paper introduces a new constraint-free concave dual formulation for the Wasserstein barycenter. Tailoring the vanilla dual gradient ascent algorithm to the Sobolev geometry, w…
UniCode: Learning a Unified Codebook for Multimodal Large Language Models
Sipeng Zheng, Bohan Zhou, Yicheng Feng +2
In this paper, we propose \textbf{UniCode}, a novel approach within the domain of multimodal large language models (MLLMs) that learns a unified codebook to efficiently tokenize vi…
NOLO: Navigate Only Look Once
Bohan Zhou, Zhongbin Zhang, Jiangxing Wang +1
The in-context learning ability of Transformer models has brought new possibilities to visual navigation. In this paper, we focus on the video navigation setting, where an in-conte…
A note on equivalences between various mixing scales
Bohan Zhou
In this note, we provide with a simple example to show a defect in the definition of the geometric mixing scale, and then introduce an improved scale, called as the strong geometri…