Publications (9)
Using Chatbots to Teach Languages
Yu Li, Chun-Yen Chen, Dian Yu +6
This paper reports on progress towards building an online language learning tool to provide learners with conversational experience by using dialog systems as conversation practice…
ErAConD : Error Annotated Conversational Dialog Dataset for Grammatical Error Correction
Xun Yuan, Derek Pham, Sam Davidson +1
Currently available grammatical error correction (GEC) datasets are compiled using well-formed written text, limiting the applicability of these datasets to other domains such as i…
Multi-IaC-Eval: Benchmarking Cloud Infrastructure as Code Across Multiple Formats
Sam Davidson, Li Sun, Bhavana Bhasker +2
Infrastructure as Code (IaC) is fundamental to modern cloud computing, enabling teams to define and manage infrastructure through machine-readable configuration files. However, dif…
Gunrock 2.0: A User Adaptive Social Conversational System
Kaihui Liang, Austin Chau, Yu Li +8
Gunrock 2.0 is built on top of Gunrock with an emphasis on user adaptation. Gunrock 2.0 combines various neural natural language understanding modules, including named entity detec…
IdEALS: Idiomatic Expressions for Advancement of Language Skills
Narutatsu Ri, Bill Sun, Sam Davidson +1
Although significant progress has been made in developing methods for Grammatical Error Correction (GEC), addressing word choice improvements has been notably lacking and enhancing…
Gunrock: A Social Bot for Complex and Engaging Long Conversations
Dian Yu, Michelle Cohn, Yi Mang Yang +12
Gunrock is the winner of the 2018 Amazon Alexa Prize, as evaluated by coherence and engagement from both real users and Amazon-selected expert conversationalists. We focus on under…
Dependency Parsing for Spoken Dialog Systems
Sam Davidson, Dian Yu, Zhou Yu
Dependency parsing of conversational input can play an important role in language understanding for dialog systems by identifying the relationships between entities extracted from…
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
Prithwish Jana, Sam Davidson, Bhavana Bhasker +3
Automating Infrastructure-as-Code (IaC) is challenging, and large language models (LLMs) often produce incorrect configurations from natural language (NL). We present TerraFormer,…
User Simulation with Large Language Models for Evaluating Task-Oriented Dialogue
Sam Davidson, Salvatore Romeo, Raphael Shu +4
One of the major impediments to the development of new task-oriented dialogue (TOD) systems is the need for human evaluation at multiple stages and iterations of the development pr…