1 citations · 1 across the 6 of their papers we have counts for
7 papers
GuardianBench: A Same-Scene Instruction-Contrastive Benchmark for Latent Contextual Risk in Embodied AI
Zhesheng Zhang, Jiahao Lu, Wei Liu +8
In embodied AI, safety risk can be latent: a benign instruction and a safe scene become hazardous only when composed. Prior work has advanced embodied safety by varying visual cont…
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
Kun Yu, Jianhua Yang, Yixiang Chen +7
Language-guided human following is an important capability for embodied agents, but existing benchmarks typically assume that the target person is visible at the start of an episod…
FlowWAM: Optical Flow as a Unified Action Representation for World Action Models
Yixiang Chen, Peiyan Li, Yuan Xu +13
World Action Models (WAMs) are able to leverage pretrained video generators for both world modeling and action prediction. However, directly leveraging such video generators for co…
YOWO-Plus: An Incremental Improvement
Jianhua Yang
In this technical report, we would like to introduce our updates to YOWO, a real-time method for spatio-temporal action detection. We make a bunch of little design changes to make…
JPEG Steganography with Embedding Cost Learning and Side-Information Estimation
Jianhua Yang, Yi Liao, Fei Shang +2
A great challenge to steganography has arisen with the wide application of steganalysis methods based on convolutional neural networks (CNNs). To this end, embedding cost learning…
CMF: Cascaded Multi-model Fusion for Referring Image Segmentation
Jianhua Yang, Yan Huang, Zhanyu Ma +1
In this work, we address the task of referring image segmentation (RIS), which aims at predicting a segmentation mask for the object described by a natural language expression. Mos…