activity
20182021
most citedParallel Corpus Filtering via Pre-trained Language Models

7 citations · 12 across the 3 of their papers we have counts for

collaborators

8 papers

cs.CL2021

MeetDot: Videoconferencing with Live Translation Captions

Arkady Arkhangorodsky, Christopher Chu, Scot Fang +5

We present MeetDot, a videoconferencing system with live translation captions overlaid on screen. The system aims to facilitate conversation between people who speak different lang…

cs.CL20215 cited

A Hybrid Task-Oriented Dialog System with Domain and Task Adaptive Pretraining

Boliang Zhang, Ying Lyu, Ning Ding +4

This paper describes our submission for the End-to-end Multi-domain Task Completion Dialog shared task at the 9th Dialog System Technology Challenge (DSTC-9). Participants in the s…

cs.CL2020

Global Attention for Name Tagging

Boliang Zhang, Spencer Whitehead, Lifu Huang +1

Many name tagging approaches use local contextual information with much success, but fail when the local context is ambiguous or limited. We present a new framework to improve name…

cs.CL2020

MEEP: An Open-Source Platform for Human-Human Dialog Collection and End-to-End Agent Training

Arkady Arkhangorodsky, Amittai Axelrod, Christopher Chu +6

We create a new task-oriented dialog platform (MEEP) where agents are given considerable freedom in terms of utterances and API calls, but are constrained to work within a push-but…

cs.CL20207 cited

Parallel Corpus Filtering via Pre-trained Language Models

Boliang Zhang, Ajay Nagesh, Kevin Knight

Web-crawled data provides a good source of parallel corpora for training machine translation models. It is automatically obtained, but extremely noisy, and recent work shows that n…

cs.CL2018

Describing a Knowledge Base

Qingyun Wang, Xiaoman Pan, Lifu Huang +4

We aim to automatically generate natural language descriptions about an input structured knowledge base (KB). We build our generation framework based on a pointer network which can…