3 papers
cs.CL2025
AutoMedic: An Automated Evaluation Framework for Clinical Conversational Agents with Medical Dataset Grounding
Gyutaek Oh, Sangjoon Park, Byung-Hoon Kim
Evaluating large language models (LLMs) has recently emerged as a critical issue for safe and trustworthy application of LLMs in the medical domain. Although a variety of static me…
cs.LG2025
Deep operator network for surrogate modeling of poroelasticity with random permeability fields
Sangjoon Park, Yeonjong Shin, Jinhyun Choo
Poroelasticity -- coupled fluid flow and elastic deformation in porous media -- often involves spatially variable permeability, especially in subsurface systems. In such cases, sim…
cs.CL2025
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
Gyutaek Oh, Seoyeon Kim, Sangjoon Park +1
Test-time scaling has recently emerged as a promising approach for enhancing the reasoning capabilities of large language models or vision-language models during inference. Althoug…