Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation
Mingfei Han, Liang Ma, Kamila Zhumakhanova +5
Vision-and-Language Navigation (VLN) suffers from the limited diversity and scale of training data, primarily constrained by the manual curation of existing simulators. To address…
cs.CV2024
Meta-Exploiting Frequency Prior for Cross-Domain Few-Shot Learning
Fei Zhou, Peng Wang, Lei Zhang +5
Meta-learning offers a promising avenue for few-shot learning (FSL), enabling models to glean a generalizable feature embedding through episodic training on synthetic FSL tasks in…