3 papers
cs.CV2026
InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model
InSpatio Team, Donghui Shen, Guofeng Zhang +16
We present InSpatio-WorldFM, an open-source real-time frame model for spatial intelligence. Unlike video-based world models that rely on sequential frame generation and incur subst…
cs.CV2023
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
Mengyuan Liu, Hong Liu, Tianyu Guo
Considering the instance-level discriminative ability, contrastive learning methods, including MoCo and SimCLR, have been adapted from the original image representation learning ta…
cs.CV2023
FSAR: Federated Skeleton-based Action Recognition with Adaptive Topology Structure and Knowledge Distillation
Jingwen Guo, Hong Liu, Shitong Sun +3
Existing skeleton-based action recognition methods typically follow a centralized learning paradigm, which can pose privacy concerns when exposing human-related videos. Federated L…