2 papers
cs.SE2026
FaultLens: Learning Compact Behavioral Test Suites for Generated Operational Programs
Zeming Liu, Hang Lyu, Jingtao Zhang
Generated operational programs are often validated with either a few hand-written examples or exhaustive regression suites. The former can miss sparse boundary and interaction faul…
cs.AI2026
Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design
Zejun Liu, Jian Wu, Ru Peng +4
AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research process, prior work has focus…