14 papers
Cell as Point: One-Stage Framework for Efficient Cell Tracking
Yaxuan Song, Jianan Fan, Heng Huang +2
Conventional multi-stage cell tracking approaches rely heavily on detection or segmentation in each frame as a prerequisite, requiring substantial resources for high-quality segmen…
Towards Conditional Feature Alignment for Cross-Domain Counting
Zhuonan Liang, Dongnan Liu, Jianan Fan +6
Object counting models often degrade under cross-domain deployment because density composition varies across domains and is itself task-relevant. Standard feature alignment methods…
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Haiwen Diao, Penghao Wu, Hanming Deng +55
Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as distinct problems, leading to fra…
RNA-FM: Flow-Matching Generative Model for Genome-wide RNA-Seq Prediction
Yaxuan Song, Jianan Fan, Tianyi Wang +4
Histopathology whole-slide images (WSIs) are routinely acquired in clinical practice and contain rich tissue morphology but lack direct molecular architecture and functional progra…
Beyond Benchmarks of IUGC: Rethinking Requirements of Deep Learning Methods for Intrapartum Ultrasound Biometry from Fetal Ultrasound Videos
Jieyun Bai, Zihao Zhou, Yitong Tang +60
A substantial proportion (45\%) of maternal deaths, neonatal deaths, and stillbirths occur during the intrapartum phase, with a particularly high burden in low- and middle-income c…
SurgiATM: A Physics-Guided Plug-and-Play Model for Deep Learning-Based Smoke Removal in Laparoscopic Surgery
Mingyu Sheng, Jianan Fan, Dongnan Liu +3
During laparoscopic surgery, smoke generated by tissue cauterization degrade endoscopic frames quality, increasing surgical risk and hindering both clinical decision-making and com…