6 papers
Does Bigger Mean Better? Comparitive Analysis of CNNs and Biomedical Vision Language Modles in Medical Diagnosis
Ran Tong, Jiaqi Liu, Tong Wang +4
The accurate interpretation of chest radiographs using automated methods is a critical task in medical imaging. This paper presents a comparative analysis between a supervised ligh…
Lightweight Baselines for Medical Abstract Classification: DistilBERT with Cross-Entropy as a Strong Default
Jiaqi Liu, Tong Wang, Su Liu +4
The research evaluates lightweight medical abstract classification methods to establish their maximum performance capabilities under financial budget restrictions. On the public me…
Renaissance of RNNs in Streaming Clinical Time Series: Compact Recurrence Remains Competitive with Transformers
Ran Tong, Jiaqi Liu, Su Liu +2
We present a compact, strictly causal benchmark for streaming clinical time series on the MIT--BIH Arrhythmia Database using per-second heart rate. Two tasks are studied under reco…
Toward Causal-Visual Programming: Enhancing Agentic Reasoning in Low-Code Environments
Jiexi Xu, Jiaqi Liu, Lanruo Wang +1
Large language model (LLM) agents are increasingly capable of orchestrating complex tasks in low-code environments. However, these agents often exhibit hallucinations and logical i…
Rainbow Noise: Stress-Testing Multimodal Harmful-Meme Detectors on LGBTQ Content
Ran Tong, Songtao Wei, Jiaqi Liu +1
Hateful memes aimed at LGBTQ\,+ communities often evade detection by tweaking either the caption, the image, or both. We build the first robustness benchmark for this setting, pair…
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes
Jiaqi Liu, Ran Tong, Aowei Shen +3
Memes often merge visuals with brief text to share humor or opinions, yet some memes contain harmful messages such as hate speech. In this paper, we introduces MemeBLIP2, a light w…