4 papers
SCARV: Structure-Constrained Aggregation for Stable Sample Ranking in Redundant NLP Datasets
Xu Zheng, Feiyu Wu, Linhong Wu +2
Sample-level rankings are increasingly used in data-centric NLP for analysis, filtering, debugging, and curation, yet existing pipelines typically score training examples pointwise…
ISEE: Interactive Semantic Enrichment for Database Fields
Yuan Tian, Yiru Chen, Rakesh R. Menon +8
LLM-based agents are increasingly being deployed for data-related tasks, including data sense-making, exploration, and retrieval. However, their performance heavily depends on the…
Text-to-SQL Domain Adaptation via Human-LLM Collaborative Data Annotation
Yuan Tian, Daniel Lee, Fei Wu +5
Text-to-SQL models, which parse natural language (NL) questions to executable SQL queries, are increasingly adopted in real-world applications. However, deploying such models in th…
Adobe Summit Concierge Evaluation with Human in the Loop
Yiru Chen, Sally Fang, Sai Sree Harsha +6
Generative AI assistants offer significant potential to enhance productivity, streamline information access, and improve user experience in enterprise contexts. In this work, we pr…