1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2025
MANZANO: A Simple and Scalable Unified Multimodal Model with a Hybrid Vision Tokenizer
Yanghao Li, Rui Qian, Bowen Pan +24
Unified multimodal Large Language Models (LLMs) that can both understand and generate visual content hold immense potential. However, existing open-source models often suffer from…
cs.CV2024
Efficient Semantic Splatting for Remote Sensing Multi-view Segmentation
Zipeng Qi, Hao Chen, Haotian Zhang +2
In this paper, we propose a novel semantic splatting approach based on Gaussian Splatting to achieve efficient and low-latency. Our method projects the RGB attributes and semantic…
cs.AI2024★ 1 cited
Scheduling Drone and Mobile Charger via Hybrid-Action Deep Reinforcement Learning
Jizhe Dou, Haotian Zhang, Guodong Sun
Recently there has been a growing interest in industry and academia, regarding the use of wireless chargers to prolong the operational longevity of unmanned aerial vehicles (common…