2 papers
cs.CV2026
TRNet: Topography-Guided Frequency Rectification and Structure-Aware Decoding for Multimodal Paddy Rice Segmentation
Kaiwen Xiao, Chunlong Fu, Liping Zheng +1
Mapping paddy rice from very-high-resolution imagery in mountainous and hilly regions is difficult because terrain alters optical appearance and increases confusion with visually s…
cs.CV2026
Singpath-VL Technical Report
Zhen Qiu, Kaiwen Xiao, Zhengwei Lu +3
We present Singpath-VL, a vision-language large model, to fill the vacancy of AI assistant in cervical cytology. Recent advances in multi-modal large language models (MLLMs) have s…