From the 1 of 2 linked papers with an AI index.
2 papers
cs.SE2026
RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists
Yuexi Yang, Alyssa Wu, Ji Luo +4
The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale assistance. However, existing ben…
cs.SE2026
Not as Sweet by Another Name: An Empirical Study of Format Robustness in LLM Document Workflows
Xiaoyu Zhang, Xianyun Cheng, Tianlin Li +3
The paper investigates how changing document formats (e.g., CSV vs. plain text) affects the reliability of end-to-end LLM-driven software workflows, introducing a metamorphic testi…