most citedHOI-aware Adaptive Network for Weakly-supervised Action Segmentation

7 citations

10 papers

cs.CV2026

IntentQA: Intent Question Answering in Videos by Cognitive Context Reasoning

Jiapeng Li, Ping Wei, Wenjuan Han +2

Video understanding requires intelligent agents to transcend mere recognition of visual facts and comprehend the underlying intents behind human actions (often termed the "dark mat…

cs.LG2026

Domain Knowledge Based Temporal-Spatial Graph Convolution Network for ECG Recognition

Wenting Ma, Zhipeng Zhang, Xiaohang Yuan +6

In light of strides in Arti cial Intelligence (AI) and its wide spread application, challenges persist in the interpretability of AI models, particularly within specialized domains…

cs.LG20261 cited

From Pixels to Temporal Correlations: Learning Informative Representations for Reinforcement Learning Pre-training

Jinwen Wang, Youfang Lin, Xiaobo Hu +4

Unsupervised pre-training on large-scale datasets has demonstrated significant potential for improving the sample efficiency and performance of Reinforcement Learning (RL). Given t…

cs.LG2026

Task-Relevant Representation Decoupling for Visual Reinforcement Learning Generalization

Jinwen Wang, Youfang Lin, Xiaobo Hu +4

Visual Reinforcement Learning (VRL) has achieved considerable success in solving control tasks. However, generalizing learned policies to new environments remains a major challenge…

eess.IV2026

Symmetric Entropy-Constrained Video Coding for Machines

Yuxiao Sun, Meiqin Liu, Chao Yao +5

As video transmission increasingly serves machine vision systems (MVS) instead of human vision systems (HVS), video coding for machines (VCM) has become a critical research topic.…

cs.LG2026

Tracking Large-scale Shared Bikes with Inertial Motion Learning in GNSS Blocked Environments

Feng Liu, Kejia Li, Zhiwei Yang +5

Although Global Navigation Satellite Systems (GNSS) provide a general solution for bike tracking outdoors, there still exist complex riding environments where only inertial navigat…