3 citations · 3 across the 4 of their papers we have counts for
5 papers
Omni Interaction Agent Technical Report
Orantqing, Shengpeng Ji, Junlong Tong +19
In this work, we present Gander, an end-to-end model that unifies omni perception, realtime interaction, and agentic capabilities within a single framework. In contrast to turn-bas…
ViCoStream: Streaming VideoLLMs Can Run Beyond 100 FPS with Stage-Wise Coordinated Inference
Yang Tan, Junlong Tong, Linan Yue +3
Streaming VideoLLMs must continuously process incoming video while maintaining low query latency, making both video-ingestion throughput and query-time responsiveness critical for…
AdaSR: Adaptive Streaming Reasoning with Hierarchical Relative Policy Optimization
Junlong Tong, Wenqi Xu, Yingqi Fan +4
Large reasoning models typically follow a read-then-think paradigm: they observe the complete input, reason over a static context, and then produce the answer. Yet many real-world…
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
Liping Xie, Yang Tan, Shicheng Jing +2
As a critical task in video sequence classification within computer vision, Online Action Detection (OAD) has garnered significant attention. The sensitivity of mainstream OAD mode…
MALT: Multi-scale Action Learning Transformer for Online Action Detection
Zhipeng Yang, Ruoyu Wang, Yang Tan +1
Online action detection (OAD) aims to identify ongoing actions from streaming video in real-time, without access to future frames. Since these actions manifest at varying scales of…