papers

Publications (198)

cs.GR2026

HoloPathTracer: Fast and Accurate Wave Path Tracing for Holography

Wenbin Zhou, Xiangyu Meng, Jiankai Xing +3

cs.CL2025

Granary: Speech Recognition and Translation Dataset in 25 European Languages

Nithin Rao Koluguri, Monica Sekoyan, George Zelenfroynd +12

cs.LG2024

Deep learning with noisy labels in medical prediction problems: a scoping review

Yishu Wei, Yu Deng, Cong Sun +3

eess.AS2023

Tensor decomposition for minimization of E2E SLU model toward on-device processing

Yosuke Kashiwagi, Siddhant Arora, Hayato Futami +6

cs.CL2023

Structured Pruning of Self-Supervised Pre-trained Models for Speech Recognition and Understanding

Yifan Peng, Kwangyoun Kim, Felix Wu +2

cs.CV2019

Holistic and Comprehensive Annotation of Clinically Significant Findings on Diverse CT Images: Learning from Radiology Reports and Label Ontology

Ke Yan, Yifan Peng, Veit Sandfort +3

cs.CV2025

A Disease-Aware Dual-Stage Framework for Chest X-ray Report Generation

Puzhen Wu, Hexin Dong, Yi Lin +2

cs.CV2021

Using Radiomics as Prior Knowledge for Thorax Disease Classification and Localization in Chest X-rays

Yan Han, Chongyan Chen, Liyan Tang +7

cs.CL2026

RSNA Large Language Model Benchmark Dataset for Chest Radiographs of Cardiothoracic Disease: Radiologist Evaluation and Validation Enhanced by AI Labels (REVEAL-CXR)

Yishu Wei, Adam E. Flanders, Errol Colak +35

eess.IV2024

Learned Scanpaths Aid Blind Panoramic Video Quality Assessment

Kanglong Fan, Wen Wen, Mu Li +2

cs.CV2026

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends

Yihao Ding, Siwen Luo, Yue Dai +6

cs.CV2024

Point Resampling and Ray Transformation Aid to Editable NeRF Models

Zhenyang Li, Zilong Chen, Feifan Qu +4

cs.CL2023

Classifying Crime Types using Judgment Documents from Social Media

Haoxuan Xu, Zeyu He, Mengfan Shen +3

eess.IV2019

A deep learning approach for automated detection of geographic atrophy from color fundus photographs

Tiarnan D. Keenan, Shazia Dharssi, Yifan Peng +5

cs.CV2024

GSCo: Towards Generalizable AI in Medicine via Generalist-Specialist Collaboration

Sunan He, Yuxiang Nie, Hongmei Wang +21

cs.CV2024

GO-NeRF: Generating Objects in Neural Radiance Fields for Virtual Reality Content Creation

Peng Dai, Feitong Tan, Xin Yu +3

cs.AI2026

Towards end-to-end LLM-based censoring-aware survival analysis

Yishu Wei, Hexin Dong, Yi Lin +3

eess.AS2024

Contextualized Automatic Speech Recognition with Attention-Based Bias Phrase Boosted Beam Search

Yui Sudo, Muhammad Shakeel, Yosuke Fukumoto +2

cs.DB2026

Efficient Vector Search in the Wild: One Model for Multi-K Queries

Yifan Peng, Jiafei Fan, Xingda Wei +7

cs.CL2025

A foundation model for human-AI collaboration in medical literature mining

Zifeng Wang, Lang Cao, Qiao Jin +20

cs.CL2024

SDoH-GPT: Using Large Language Models to Extract Social Determinants of Health (SDoH)

Bernardo Consoli, Xizhi Wu, Song Wang +10

cs.RO2026

A High-accuracy Event-based Underwater SLAM System

Yifan Peng, Qihang Liu, Haoying Li +3

eess.IV2021

Lymph Node Detection in T2 MRI with Transformers

Tejas Sudharshan Mathai, Sungwon Lee, Daniel C. Elton +4

eess.IV2023

High-performance Data Management for Whole Slide Image Analysis in Digital Pathology

Haoju Leng, Ruining Deng, Shunxing Bao +8

cs.CL2022

A Study on the Integration of Pre-trained SSL, ASR, LM and SLU Models for Spoken Language Understanding

Yifan Peng, Siddhant Arora, Yosuke Higuchi +6

cs.CV2026

PhyGaP: Physically-Grounded Gaussians with Polarization Cues

Jiale Wu, Xiaoyang Bai, Zongqi He +2

cs.AI2024

A survey of recent methods for addressing AI fairness and bias in biomedicine

Yifan Yang, Mingquan Lin, Han Zhao +3

cs.CL2025

OpusLM: A Family of Open Unified Speech Language Models

Jinchuan Tian, William Chen, Yifan Peng +9

cs.AI2026

MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation

Yi Lin, Yihao Ding, Yonghui Wu +1

cs.CL2023

A Comparative Study on E-Branchformer vs Conformer in Speech Recognition, Translation, and Understanding Tasks

Yifan Peng, Kwangyoun Kim, Felix Wu +7

cs.CL2023

A Study on the Integration of Pipeline and E2E SLU systems for Spoken Semantic Parsing toward STOP Quality Challenge

Siddhant Arora, Hayato Futami, Shih-Lun Wu +6

cs.LG2021

RadBERT-CL: Factually-Aware Contrastive Learning For Radiology Report Classification

Ajay Jaiswal, Liyan Tang, Meheli Ghosh +3

cs.CV2024

Towards long-tailed, multi-label disease classification from chest X-ray: Overview of the CXR-LT challenge

Gregory Holste, Yiliang Zhou, Song Wang +22

cs.CL2024

On the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models

Jinchuan Tian, Yifan Peng, William Chen +3

cs.CV2026

SAP: Segment Any 4K Panorama

Lutao Jiang, Zidong Cao, Weikai Chen +14

cs.CL2022

Classifying Cyber-Risky Clinical Notes by Employing Natural Language Processing

Suzanna Schmeelk, Martins Samuel Dogo, Yifan Peng +1

eess.IV2020

Michelson Holography: Dual-SLM Holography with Camera-in-the-loop Optimization

Suyeon Choi, Jonghyun Kim, Yifan Peng +1

eess.AS2022

SpeechLMScore: Evaluating speech generation using speech language model

Soumi Maiti, Yifan Peng, Takaaki Saeki +1

cs.AI2024

Leveraging Generative AI for Clinical Evidence Summarization Needs to Ensure Trustworthiness

Gongbo Zhang, Qiao Jin, Denis Jered McInerney +11

cs.IR2026

Improving Retrieval-Augmented Generation without Taxonomy-based Error Categorization

Gongbo Zhang, Yifan Peng, Chunhua Weng

cs.CV2017

ChestX-ray8: Hospital-scale Chest X-ray Database and Benchmarks on Weakly-Supervised Classification and Localization of Common Thorax Diseases

Xiaosong Wang, Yifan Peng, Le Lu +3

cs.CL2023

Utilizing Longitudinal Chest X-Rays and Reports to Pre-Fill Radiology Reports

Qingqing Zhu, Tejas Sudharshan Mathai, Pritam Mukherjee +3

cs.CV2019

DeepSeeNet: A deep learning model for automated classification of patient-based age-related macular degeneration severity from color fundus photographs

Yifan Peng, Shazia Dharssi, Qingyu Chen +5

cs.CV2025

SynDoc: A Hybrid Discriminative-Generative Framework for Enhancing Synthetic Domain-Adaptive Document Key Information Extraction

Yihao Ding, Soyeon Caren Han, Yanbei Jiang +3

cs.CV2026

MonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report Generation

Yi Lin, Yihao Ding, Elana Benishay +8

MonteRET is an AI system that combines whole‑volume CT features with region‑level anatomical information and knowledge retrieval to automatically generate more complete and clinica…

#chest ct#report generation#multimodal retrieval#region-aware modeling
cs.CL2024

An Empirical Study of Speech Language Models for Prompt-Conditioned Speech Synthesis

Yifan Peng, Ilia Kulikov, Yilin Yang +4

cs.CV2026

E2Pano: Learning Event-to-Panorama Image Reconstruction

Zhenyang Li, Zongqi He, Jia Pan +2

cs.CL2023

Joint Prediction and Denoising for Large-scale Multilingual Self-supervised Learning

William Chen, Jiatong Shi, Brian Yan +6

cs.CV2022

Radiomics-Guided Global-Local Transformer for Weakly Supervised Pathology Localization in Chest X-Rays

Yan Han, Gregory Holste, Ying Ding +3

eess.IV2024

Evaluating GPT-4 with Vision on Detection of Radiological Findings on Chest Radiographs

Yiliang Zhou, Hanley Ong, Patrick Kennedy +6

cs.CL2026

Comparing LLM and Fine-Tuned Model Performance on NVDRS Circumstance Extraction with Varying Prompt Complexity

Geoffrey Martin, Xuan Zhong Feng, Yifan Peng

cs.CL2024

Demonstration-based learning for few-shot biomedical named entity recognition under machine reading comprehension

Leilei Su, Jian Chen, Yifan Peng +1

cs.CL2025

Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications

Mingxuan Liu, Yuhe Ke, Wentao Zhu +9

eess.AS2024

Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks

Soumi Maiti, Yifan Peng, Shukjae Choi +3

cs.CL2025

DYNAC: Dynamic Vocabulary based Non-Autoregressive Contextualization for Speech Recognition

Yui Sudo, Yosuke Fukumoto, Muhammad Shakeel +3

cs.SD2024

ESPnet-EZ: Python-only ESPnet for Easy Fine-tuning and Integration

Masao Someki, Kwanghee Choi, Siddhant Arora +7

cs.CL2026

Budget-Aware Routing for Long Clinical Text

Khizar Qureshi, Geoffrey Martin, Yifan Peng

cs.AI2026

Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports

Yishu Wei, Yi Lin, Adam Flanders +2

cs.CL2024

DIRI: Adversarial Patient Reidentification with Large Language Models for Evaluating Clinical Text Anonymization

John X. Morris, Thomas R. Campion, Sri Laasya Nutheti +4

cs.CL2024

A Framework for Human Evaluation of Large Language Models in Healthcare Derived from Literature Review

Thomas Yu Chow Tam, Sonish Sivarajkumar, Sumit Kapoor +12

cs.CV2023

Learning a Room with the Occ-SDF Hybrid: Signed Distance Function Mingled with Occupancy Aids Scene Representation

Xiaoyang Lyu, Peng Dai, Zizhang Li +4

cs.CV2026

JuZhou 1.0 Technical Report: The First Edge-Native Text-to-Image Foundation Model Trained Entirely on China-Developed AI Accelerators

Ce Chen, Congrui Wang, Yonglin Li +23

cs.CL2024

Uncovering Misattributed Suicide Causes through Annotation Inconsistency Detection in Death Investigation Notes

Song Wang, Yiliang Zhou, Ziqiang Han +5

cs.CY2024

Environment Scan of Generative AI Infrastructure for Clinical and Translational Science

Betina Idnay, Zihan Xu, William G. Adams +54

cs.CY2021

When Text Simplification Is Not Enough: Could a Graph-Based Visualization Facilitate Consumers' Comprehension of Dietary Supplement Information?

Xing He, Rui Zhang, Jordan Alpert +7

cs.CL2026

Scalable Scientific Interest Profiling Using Large Language Models

Yilun Liang, Gongbo Zhang, Edward Sun +6

cs.CL2026

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

Qiao Jin, Yin Fang, Lauren He +12

q-bio.GN2024

Deciphering genomic codes using advanced NLP techniques: a scoping review

Shuyan Cheng, Yishu Wei, Yiliang Zhou +4

cs.CL2021

Leveraging Deep Representations of Radiology Reports in Survival Analysis for Predicting Heart Failure Patient Mortality

Hyun Gi Lee, Evan Sholle, Ashley Beecy +2

eess.AS2026

Unifying Diarization, Separation, and ASR with Multi-Speaker Encoder

Muhammad Shakeel, Yui Sudo, Yifan Peng +2

cs.CV2019

MULAN: Multitask Universal Lesion Analysis Network for Joint Lesion Detection, Tagging, and Segmentation

Ke Yan, Youbao Tang, Yifan Peng +4

cs.AI2026

Entry-level guide to the use of large language models for medical research

Qiao Jin, Nicholas Wan, Robert Leaman +20

cs.AI2025

Enhancing Health Fact-Checking with LLM-Generated Synthetic Data

Jingze Zhang, Jiahe Qian, Yiliang Zhou +1

cs.CV2025

Enhanced Velocity Field Modeling for Gaussian Video Reconstruction

Zhenyang Li, Xiaoyang Bai, Tongchen Zhang +3

cs.CV2025

Glossy Object Reconstruction with Cost-effective Polarized Acquisition

Bojian Wu, Yifan Peng, Ruizhen Hu +1

stat.ML2025

Multivariate Density Estimation via Variance-Reduced Sketching

Yifan Peng, Yuehaw Khoo, Daren Wang

cs.CL2026

Using reasoning LLMs to extract SDOH events from clinical notes

Ertan Dogan, Kunyu Yu, Yifan Peng

math.NA2025

Optimization-Free Diffusion Model -- A Perturbation Theory Approach

Yuehaw Khoo, Mathias Oster, Yifan Peng

cs.CV2025

UniBiomed: A Universal Foundation Model for Grounded Biomedical Image Interpretation

Linshan Wu, Yuxiang Nie, Sunan He +12

cs.CL2025

A Multi-Stage Large Language Model Framework for Extracting Suicide-Related Social Determinants of Health

Song Wang, Yishu Wei, Haotian Ma +10

eess.AS2022

E-Branchformer: Branchformer with Enhanced merging for speech recognition

Kwangyoun Kim, Felix Wu, Yifan Peng +4

cs.LG2026

Large Language Models Lack Temporal Awareness of Medical Knowledge

Zihan Guan, Qiao Jin, Guangzhi Xiong +6

math.NA2026

Generative Modeling via Hierarchical Tensor Sketching

Yifan Peng, Yian Chen, E. Miles Stoudenmire +1

cs.AI2025

Orchestrator Multi-Agent Clinical Decision Support System for Secondary Headache Diagnosis in Primary Care

Xizhi Wu, Nelly Estefanie Garduno-Rapp, Justin F Rousseau +6

cs.AI2026

CPGPrompt: Translating Clinical Guidelines into LLM-Executable Decision Support

Ruiqi Deng, Geoffrey Martin, Tony Wang +6

cs.CL2025

A Multi-agent Large Language Model Framework to Automatically Assess Performance of a Clinical AI Triage Tool

Adam E. Flanders, Yifan Peng, Luciano Prevedello +6

cs.CV2025

EventTracer: Fast Path Tracing-based Event Stream Rendering

Zhenyang Li, Xiaoyang Bai, Jinfan Lu +3

cs.CL2024

A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models

Gongbo Zhang, Zihan Xu, Qiao Jin +8

eess.AS2024

Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss

Muhammad Shakeel, Yui Sudo, Yifan Peng +1

cs.CY2023

From Military to Healthcare: Adopting and Expanding Ethical Principles for Generative Artificial Intelligence

David Oniani, Jordan Hilsman, Yifan Peng +5

cs.SD2025

Improving Multilingual Speech Models on ML-SUPERB 2.0: Fine-tuning with Data Augmentation and LID-Aware CTC

Qingzheng Wang, Jiancheng Sun, Yifan Peng +1

cs.CL2025

Natural Language Processing in Support of Evidence-based Medicine: A Scoping Review

Zihan Xu, Haotian Ma, Gongbo Zhang +3

cs.CL2024

Multi-Convformer: Extending Conformer with Multiple Convolution Kernels

Darshan Prabhu, Yifan Peng, Preethi Jyothi +1

cs.CV2022

Long-Tailed Classification of Thorax Diseases on Chest X-Ray: A New Benchmark Study

Gregory Holste, Song Wang, Ziyu Jiang +5

cs.CL2018

Chemical-protein relation extraction with ensembles of SVM, CNN, and RNN models

Yifan Peng, Anthony Rios, Ramakanth Kavuluru +1

cs.CL2024

OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer

Yifan Peng, Jinchuan Tian, William Chen +9

cs.CV2026

Overview of the CXR-LT 2026 Challenge: Multi-Center Long-Tailed and Zero Shot Chest X-ray Classification

Hexin Dong, Yi Lin, Pengyu Zhou +7

cs.CV2025

Two-Stage Decoupling Framework for Variable-Length Glaucoma Prognosis

Yiran Song, Yikai Zhang, Silvia Orengo-Nania +5

cs.LG2026

Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence

NVIDIA, :, Amala Sanjay Deshmukh +204

cs.CL2025

MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs

Juncheng Wu, Wenlong Deng, Xingxuan Li +12