2 papers
cs.CV2026
MAPS: Multi-Anchor Projection Similarity for Joint Vision-Language Geo-Localization
Yutong Hu, Siyuan Tan, Shaocheng Yan +3
Humans localize places by integrating perceptual cues from vision with semantic reasoning from language, forming a scene understanding that is both intuitive and structured. Althou…
cs.AI2026
Mask-Proof: An LLM-based Automated Data Curation Pipeline on Mathematical Proofs
Jierui Zhang, Siyuan Tan, Xinhang Li +8
Large language models (LLMs) are increasingly capable of mathematical problem solving and can even assist with research-level proofs, yet we still lack a scalable and reproducible…