papers

Publications (8)

cs.CV2025

PathoHR: Hierarchical Reasoning for Vision-Language Models in Pathology

Yating Huang, Ziyan Huang, Lintao Xiang +2

Accurate analysis of pathological images is essential for automated tumor diagnosis but remains challenging due to high structural similarity and subtle morphological variations in…

eess.SY2023

Lattice piecewise affine approximation of explicit nonlinear model predictive control with application to trajectory tracking of mobile robot

Kangbo Wang, Kaijie Zhang, Yating Huang +1

To promote the widespread use of mobile robots in diverse fields, the performance of trajectory tracking must be ensured. To address the constraints and nonlinear features associat…

cs.CV2025

PointGS: Point Attention-Aware Sparse View Synthesis with Gaussian Splatting

Lintao Xiang, Hongpei Zheng, Yating Huang +2

3D Gaussian splatting (3DGS) is an innovative rendering technique that surpasses the neural radiance field (NeRF) in both rendering speed and visual quality by leveraging an explic…

cs.CV2025

Interpretable Multimodal Cancer Prototyping with Whole Slide Images and Incompletely Paired Genomics

Yupei Zhang, Yating Huang, Wanming Hu +3

Multimodal approaches that integrate histology and genomics hold strong potential for precision oncology. However, phenotypic and genotypic heterogeneity limits the quality of intr…

eess.IV2025

QWD-GAN: Quality-aware Wavelet-driven GAN for Unsupervised Medical Microscopy Images Denoising

Qijun Yang, Yating Huang, Lintao Xiang +1

Image denoising plays a critical role in biomedical and microscopy imaging, especially when acquiring wide-field fluorescence-stained images. This task faces challenges in multiple…

cs.CR2025

SyncGait: Robust Long-Distance Authentication for Drone Delivery via Implicit Gait Behaviors

Zijian Ling, Man Zhou, Hongda Zhai +5

In recent years, drone delivery, which utilizes unmanned aerial vehicles (UAVs) for package delivery and pickup, has gradually emerged as a crucial method in logistics. Since deliv…

eess.AS2022

LiMuSE: Lightweight Multi-modal Speaker Extraction

Qinghua Liu, Yating Huang, Yunzhe Hao +2

Multi-modal cues, including spatial information, facial expression and voiceprint, are introduced to the speech separation and speaker extraction tasks to serve as complementary in…

cs.CV2025

ITC-RWKV: Interactive Tissue-Cell Modeling with Recurrent Key-Value Aggregation for Histopathological Subtyping

Yating Huang, Qijun Yang, Lintao Xiang +1

Accurate interpretation of histopathological images demands integration of information across spatial and semantic scales, from nuclear morphology and cellular textures to global t…