3 papers
cs.AI2026
Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems
Rebecca L. Johnson
In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: benchmarks do not just measure,…
cs.AI2026
How to Build AI Agents by Augmenting LLMs with Codified Human Expert Domain Knowledge? A Software Engineering Framework
Choro Ulan uulu, Mikhail Kulyabin, Iris Fuhrmann +6
Critical domain knowledge typically resides with few experts, creating organizational bottlenecks in scalability and decision-making. Non-experts struggle to create effective visua…
cs.HC2026
Tables or Sankey Diagrams? Investigating User Interaction with Different Representations of Simulation Parameters
Choro Ulan uulu, Mikhail Kulyabin, Katharina M Zeiner +6
Understanding complex parameter dependencies is critical for effective configuration and maintenance of software systems across diverse domains - from Computer-Aided Engineering (C…