Publications (8)
PathoHR: Hierarchical Reasoning for Vision-Language Models in Pathology
Yating Huang, Ziyan Huang, Lintao Xiang +2
Accurate analysis of pathological images is essential for automated tumor diagnosis but remains challenging due to high structural similarity and subtle morphological variations in…
Lattice piecewise affine approximation of explicit nonlinear model predictive control with application to trajectory tracking of mobile robot
Kangbo Wang, Kaijie Zhang, Yating Huang +1
To promote the widespread use of mobile robots in diverse fields, the performance of trajectory tracking must be ensured. To address the constraints and nonlinear features associat…
PointGS: Point Attention-Aware Sparse View Synthesis with Gaussian Splatting
Lintao Xiang, Hongpei Zheng, Yating Huang +2
3D Gaussian splatting (3DGS) is an innovative rendering technique that surpasses the neural radiance field (NeRF) in both rendering speed and visual quality by leveraging an explic…
Interpretable Multimodal Cancer Prototyping with Whole Slide Images and Incompletely Paired Genomics
Yupei Zhang, Yating Huang, Wanming Hu +3
Multimodal approaches that integrate histology and genomics hold strong potential for precision oncology. However, phenotypic and genotypic heterogeneity limits the quality of intr…
QWD-GAN: Quality-aware Wavelet-driven GAN for Unsupervised Medical Microscopy Images Denoising
Qijun Yang, Yating Huang, Lintao Xiang +1
Image denoising plays a critical role in biomedical and microscopy imaging, especially when acquiring wide-field fluorescence-stained images. This task faces challenges in multiple…
SyncGait: Robust Long-Distance Authentication for Drone Delivery via Implicit Gait Behaviors
Zijian Ling, Man Zhou, Hongda Zhai +5
In recent years, drone delivery, which utilizes unmanned aerial vehicles (UAVs) for package delivery and pickup, has gradually emerged as a crucial method in logistics. Since deliv…
LiMuSE: Lightweight Multi-modal Speaker Extraction
Qinghua Liu, Yating Huang, Yunzhe Hao +2
Multi-modal cues, including spatial information, facial expression and voiceprint, are introduced to the speech separation and speaker extraction tasks to serve as complementary in…
ITC-RWKV: Interactive Tissue-Cell Modeling with Recurrent Key-Value Aggregation for Histopathological Subtyping
Yating Huang, Qijun Yang, Lintao Xiang +1
Accurate interpretation of histopathological images demands integration of information across spatial and semantic scales, from nuclear morphology and cellular textures to global t…