2 papers
cs.HC2024
A Framework for Evaluating Appropriateness, Trustworthiness, and Safety in Mental Wellness AI Chatbots
Lucia Chen, David A. Preece, Pilleriin Sikka +2
Large language model (LLM) chatbots are susceptible to biases and hallucinations, but current evaluations of mental wellness technologies lack comprehensive case studies to evaluat…
cs.CL2024
AutoGRAMS: Autonomous Graphical Agent Modeling Software
Ben Krause, Lucia Chen, Emmanuel Kahembwe
We introduce the AutoGRAMS framework for programming multi-step interactions with language models. AutoGRAMS represents AI agents as a graph, where each node can execute either a l…