Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Geo-Embed: Towards Unified Multimodal Embeddings for Urban Understanding
Jiapeng Li, Yong Li, Junjie Zhou +2
Geospatial and urban applications increasingly require models to compare heterogeneous evidence across street-view imagery, remote-sensing observations, text descriptions, region p…
cs.CV2025
Wan: Open and Advanced Large-Scale Video Generative Models
Team Wan, Ang Wang, Baole Ai +58
This report presents Wan, a comprehensive and open suite of video foundation models designed to push the boundaries of video generation. Built upon the mainstream diffusion transfo…