Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment
Ali Pourghasemi Fatideh, Wilder Baldwin, Maria Dhakal +2
LLM-based dialogue assistants have become mainstream tools for software developers, yet current evaluation benchmarks focus exclusively on functional correctness. This leaves a cri…
cs.AI2026
Knowledge Graph Representations for LLM-Based Policy Compliance Reasoning
Wilder Baldwin, Sepideh Ghanavati
The risks posed by AI features are increasing as they are rapidly integrated into software applications. In response, regulations and standards for safe and secure AI have been pro…