papers

Publications (9)

cs.CL2022

Using Chatbots to Teach Languages

Yu Li, Chun-Yen Chen, Dian Yu +6

This paper reports on progress towards building an online language learning tool to provide learners with conversational experience by using dialog systems as conversation practice…

cs.CL2022

ErAConD : Error Annotated Conversational Dialog Dataset for Grammatical Error Correction

Xun Yuan, Derek Pham, Sam Davidson +1

Currently available grammatical error correction (GEC) datasets are compiled using well-formed written text, limiting the applicability of these datasets to other domains such as i…

cs.DC2025

Multi-IaC-Eval: Benchmarking Cloud Infrastructure as Code Across Multiple Formats

Sam Davidson, Li Sun, Bhavana Bhasker +2

Infrastructure as Code (IaC) is fundamental to modern cloud computing, enabling teams to define and manage infrastructure through machine-readable configuration files. However, dif…

cs.CL2020

Gunrock 2.0: A User Adaptive Social Conversational System

Kaihui Liang, Austin Chau, Yu Li +8

Gunrock 2.0 is built on top of Gunrock with an emphasis on user adaptation. Gunrock 2.0 combines various neural natural language understanding modules, including named entity detec…

cs.CL2023

IdEALS: Idiomatic Expressions for Advancement of Language Skills

Narutatsu Ri, Bill Sun, Sam Davidson +1

Although significant progress has been made in developing methods for Grammatical Error Correction (GEC), addressing word choice improvements has been notably lacking and enhancing…

cs.CL2019

Gunrock: A Social Bot for Complex and Engaging Long Conversations

Dian Yu, Michelle Cohn, Yi Mang Yang +12

Gunrock is the winner of the 2018 Amazon Alexa Prize, as evaluated by coherence and engagement from both real users and Amazon-selected expert conversationalists. We focus on under…

cs.CL2019

Dependency Parsing for Spoken Dialog Systems

Sam Davidson, Dian Yu, Zhou Yu

Dependency parsing of conversational input can play an important role in language understanding for dialog systems by identifying the relationships between entities extracted from…

cs.SE2026

TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback

Prithwish Jana, Sam Davidson, Bhavana Bhasker +3

Automating Infrastructure-as-Code (IaC) is challenging, and large language models (LLMs) often produce incorrect configurations from natural language (NL). We present TerraFormer,…

cs.CL2023

User Simulation with Large Language Models for Evaluating Task-Oriented Dialogue

Sam Davidson, Salvatore Romeo, Raphael Shu +4

One of the major impediments to the development of new task-oriented dialogue (TOD) systems is the need for human evaluation at multiple stages and iterations of the development pr…