3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.AI2026
LiveMedBench: A Contamination-Free Medical Benchmark for LLMs with Automated Rubric Evaluation
Zhiling Yan, Dingjie Song, Zhe Fang +4
The deployment of Large Language Models (LLMs) in high-stakes clinical settings demands rigorous and reliable evaluation. However, existing medical benchmarks remain static, suffer…
eess.IV2024
TTT-Unet: Enhancing U-Net with Test-Time Training Layers for Biomedical Image Segmentation
Rong Zhou, Zhengqing Yuan, Zhiling Yan +7
Biomedical image segmentation is crucial for accurately diagnosing and analyzing various diseases. However, Convolutional Neural Networks (CNNs) and Transformers, the most commonly…
cs.CL2024★ 3 cited
Can Large Language Models Automatically Jailbreak GPT-4V?
Yuanwei Wu, Yue Huang, Yixin Liu +3
GPT-4V has attracted considerable attention due to its extraordinary capacity for integrating and processing multimodal information. At the same time, its ability of face recogniti…