4 papers
Bias in the Loop: Auditing LLM-as-a-Judge for Software Engineering
Zixiao Zhao, Amirreza Esmaeili, Fatemeh Fard
Large Language Models are increasingly used as judges to evaluate code artifacts when exhaustive human review or executable test coverage is unavailable. LLM-judge is increasingly…
AstroVLM: Expert Multi-agent Collaborative Reasoning for Astronomical Imaging Quality Diagnosis
Yaohui Han, Tianshuo Wang, Zixi Zhao +6
Vision Language Models (VLMs) have been applied to several specific domains and have shown strong problem-solving capabilities. However, astronomical imaging, a quite complex probl…
Do Current Language Models Support Code Intelligence for R Programming Language?
ZiXiao Zhao, Fatemeh H. Fard
Recent advancements in developing Pre-trained Language Models for Code (Code-PLMs) have urged many areas of Software Engineering (SE) and brought breakthrough results for many SE t…
Studying Vulnerable Code Entities in R
Zixiao Zhao, Millon Madhur Das, Fatemeh H. Fard
Pre-trained Code Language Models (Code-PLMs) have shown many advancements and achieved state-of-the-art results for many software engineering tasks in the past few years. These mod…