From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Explaining Process Control Optimisation Recommendations via GradientSHAP and Implicit Differentiation
Paul Darm, Cem Alpturk, Kenneth Ulrich +3
The paper proposes a method that combines implicit differentiation with GradientSHAP and large language models to generate fast, real-time explanations for industrial process contr…
cs.CL2025
Head-Specific Intervention Can Induce Misaligned AI Coordination in Large Language Models
Paul Darm, Annalisa Riccardi
Robust alignment guardrails for large language models (LLMs) are becoming increasingly important with their widespread application. In contrast to previous studies, we demonstrate…
cs.AI2025
Inference-Time Intervention in Large Language Models for Reliable Requirement Verification
Paul Darm, James Xie, Annalisa Riccardi
Steering the behavior of Large Language Models (LLMs) remains a challenge, particularly in engineering applications where precision and reliability are critical. While fine-tuning…