activity
20182021
most citedMeta-SAC: Auto-tune the Entropy Temperature of Soft Actor-Critic via Metagradient

18 citations · 23 across the 3 of their papers we have counts for

collaborators

5 papers

cs.RO20211 cited

Adaptive Agent Architecture for Real-time Human-Agent Teaming

Tianwei Ni, Huao Li, Siddharth Agrawal +6

Teamwork is a set of interrelated reasoning, actions and behaviors of team members that facilitate common objectives. Teamwork theory and experiments have resulted in a set of stat…

cs.LG20204 cited

f-IRL: Inverse Reinforcement Learning via State Marginal Matching

Tianwei Ni, Harshit Sikchi, Yufei Wang +3

Imitation learning is well-suited for robotic tasks where it is difficult to directly program the behavior or specify a cost for optimal control. In this work, we propose a method…

cs.LG202018 cited

Meta-SAC: Auto-tune the Entropy Temperature of Soft Actor-Critic via Metagradient

Yufei Wang, Tianwei Ni

Exploration-exploitation dilemma has long been a crucial issue in reinforcement learning. In this paper, we propose a new approach to automatically balance between these two. Our m…

cs.CV2018

Elastic Boundary Projection for 3D Medical Image Segmentation

Tianwei Ni, Lingxi Xie, Huangjie Zheng +2

We focus on an important yet challenging problem: using a 2D deep network to deal with 3D segmentation for medical image analysis. Existing approaches either applied multi-view pla…

cs.CV2018

Phase Collaborative Network for Two-Phase Medical Image Segmentation

Huangjie Zheng, Lingxi Xie, Tianwei Ni +5

In real-world practice, medical images acquired in different phases possess complementary information, {\em e.g.}, radiologists often refer to both arterial and venous scans in ord…