papers

Publications (15)

cs.CL2025

Enhancing Medical Dialogue Generation through Knowledge Refinement and Dynamic Prompt Adjustment

Hongda Sun, Jiaren Peng, Wenzhong Yang +3

Medical dialogue systems (MDS) have emerged as crucial online platforms for enabling multi-turn, context-aware conversations with patients. However, existing MDS often struggle to…

cs.AI2023

Collaborative Synthesis of Patient Records through Multi-Visit Health State Inference

Hongda Sun, Hongzhan Lin, Rui Yan

Electronic health records (EHRs) have become the foundation of machine learning applications in healthcare, while the utility of real patient records is often limited by privacy an…

cs.CV2026

Geoparsing: Diagram Parsing for Plane and Solid Geometry with a Unified Formal Language

Peijie Wang, Ming-Liang Zhang, Jun Cao +10

Multimodal Large Language Models (MLLMs) have achieved remarkable progress but continue to struggle with geometric reasoning, primarily due to the perception bottleneck regarding f…

q-bio.QM2024

MIN: Multi-channel Interaction Network for Drug-Target Interaction with Protein Distillation

Shuqi Li, Shufang Xie, Hongda Sun +4

Traditional drug discovery processes are both time-consuming and require extensive professional expertise. With the accumulation of drug-target interaction (DTI) data from experime…

cs.MA2025

MobileSteward: Integrating Multiple App-Oriented Agents with Self-Evolution to Automate Cross-App Instructions

Yuxuan Liu, Hongda Sun, Wei Liu +3

Mobile phone agents can assist people in automating daily tasks on their phones, which have emerged as a pivotal research spotlight. However, existing procedure-oriented agents str…

cs.CV2026

GeoEyes: On-Demand Visual Focusing for Evidence-Grounded Understanding of Ultra-High-Resolution Remote Sensing Imagery

Fengxiang Wang, Mingshuo Chen, Yueying Li +10

The "thinking-with-images" paradigm enables multimodal large language models (MLLMs) to actively explore visual scenes via zoom-in tools. This is essential for ultra-high-resolutio…