3 papers
cs.SE2026
LLM-as-a-Judge for Human-AI Co-Creation: A Reliability-Aware Evaluation Framework for Coding
Md Faizul Ibne Amin, Yutaka Watanobe, Daniel M. Muepu +3
LLMs are increasingly employed both as judges for evaluating open-ended outputs and as co-creation partners in AI-assisted programming; yet rigorous evaluation in human-AI co-creat…
cs.SE2026
Error Understanding in Program Code: A Systematic Study of LLM-DL Combinations for Multi-label Classification
Md Faizul Ibne Amin, Yutaka Watanobe, Md. Mostafizer Rahman +2
Programming is a core skill in CS and SE, yet identifying and resolving code errors remains challenging for practitioners. LLMs have shown remarkable capabilities in NL understandi…
cs.SE2026
CodeT5-RNN: Reinforcing Contextual Embeddings for Enhanced Code Comprehension
Md Mostafizer Rahman, Ariful Islam Shiplu, Yutaka Watanobe +3
Contextual embeddings generated by LLMs exhibit strong positional inductive biases, which can limit their ability to fully capture long-range, order-sensitive dependencies in highl…