1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.LG2024
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
Sanjiban Choudhury, Paloma Sodhi
While large language models (LLMs) show impressive decision-making abilities, current methods lack a mechanism for automatic self-improvement from errors during task execution. We…
cs.CL2023★ 1 cited
Workflow-Guided Response Generation for Task-Oriented Dialogue
Do June Min, Paloma Sodhi, Ramya Ramakrishnan
Task-oriented dialogue (TOD) systems aim to achieve specific goals through interactive dialogue. Such tasks usually involve following specific workflows, i.e. executing a sequence…
cs.LG2023
SteP: Stacked LLM Policies for Web Actions
Paloma Sodhi, S. R. K. Branavan, Yoav Artzi +1
Performing tasks on the web presents fundamental challenges to large language models (LLMs), including combinatorially large open-world tasks and variations across web interfaces.…