5 citations · 14 across the 22 of their papers we have counts for
8 papers · 1 filter
SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs
Shail Desai, Aditya Pawar, Li Lin +2
Artificial Intelligence (AI) has made it possible for anyone to create images, audio, and video with unprecedented ease, enriching education, communication, and creative expression…
MedChat: A Multi-Agent Framework for Multimodal Diagnosis with Large Language Models
Philip R. Liu, Sparsh Bansal, Jimmy Dinh +6
The integration of deep learning-based glaucoma detection with large language models (LLMs) presents an automated strategy to mitigate ophthalmologist shortages and improve clinica…
Preserving AUC Fairness in Learning with Noisy Protected Groups
Mingyang Wu, Li Lin, Wenbin Zhang +3
The Area Under the ROC Curve (AUC) is a key metric for classification, especially under class imbalance, with growing research focus on optimizing AUC over accuracy in applications…
Improving Generalization of Medical Image Registration Foundation Model
Jing Hu, Kaiwei Yu, Hongjiang Xian +2
Deformable registration is a fundamental task in medical image processing, aiming to achieve precise alignment by establishing nonlinear correspondences between images. Traditional…
RLMiniStyler: Light-weight RL Style Agent for Arbitrary Sequential Neural Style Generation
Jing Hu, Chengming Feng, Shu Hu +4
Arbitrary style transfer aims to apply the style of any given artistic image to another content image. Still, existing deep learning-based methods often require significant computa…
Robust Fairness Vision-Language Learning for Medical Image Analysis
Sparsh Bansal, Mingyang Wu, Xin Wang +1
The advent of Vision-Language Models (VLMs) in medical image analysis has the potential to help process multimodal inputs and increase performance over traditional inference method…