3 papers
cs.CV2025
Enhancing Screen Time Identification in Children with a Multi-View Vision Language Model and Screen Time Tracker
Xinlong Hou, Sen Shen, Xueshen Li +7
Being able to accurately monitor the screen exposure of young children is important for research on phenomena linked to screen use such as childhood obesity, physical activity, and…
cs.AI2025
ProMRVL-CAD: Proactive Dialogue System with Multi-Round Vision-Language Interactions for Computer-Aided Diagnosis
Xueshen Li, Xinlong Hou, Ziyi Huang +1
Recent advancements in large language models (LLMs) have demonstrated extraordinary comprehension capabilities with remarkable breakthroughs on various vision-language tasks. Howev…
cs.CL2024
A Two-Stage Proactive Dialogue Generator for Efficient Clinical Information Collection Using Large Language Model
Xueshen Li, Xinlong Hou, Nirupama Ravi +2
Efficient patient-doctor interaction is among the key factors for a successful disease diagnosis. During the conversation, the doctor could query complementary diagnostic informati…