2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2026
FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-Language Navigation
Jing Zuo, Lingzhou Mu, Fan Jiang +3
Achieving human-level performance in Vision-and-Language Navigation (VLN) requires an embodied agent to jointly understand multimodal instructions and visual-spatial context while…
cs.SD2025
Magnitude-Phase Dual-Path Speech Enhancement Network based on Self-Supervised Embedding and Perceptual Contrast Stretch Boosting
Alimjan Mattursun, Liejun Wang, Yinfeng Yu +1
Speech self-supervised learning (SSL) has made great progress in various speech processing tasks, but there is still room for improvement in speech enhancement (SE). This paper pre…
cs.CV2022★ 2 cited
CrossRectify: Leveraging Disagreement for Semi-supervised Object Detection
Chengcheng Ma, Xingjia Pan, Qixiang Ye +3
Semi-supervised object detection has recently achieved substantial progress. As a mainstream solution, the self-labeling-based methods train the detector on both labeled data and u…