TabGenie: A Toolkit for Table-to-Text Generation
arXiv:2302.14169 · doi:10.18653/v1/2023.acl-demo.42
Abstract
Heterogenity of data-to-text generation datasets limits the research on data-to-text generation systems. We present TabGenie - a toolkit which enables researchers to explore, preprocess, and analyze a variety of data-to-text generation datasets through the unified framework of table-to-text generation. In TabGenie, all the inputs are represented as tables with associated metadata. The tables can be explored through the web interface, which also provides an interactive mode for debugging table-to-text generation, facilitates side-by-side comparison of generated system outputs, and allows easy exports for manual analysis. Furthermore, TabGenie is equipped with command line processing tools and Python bindings for unified dataset loading and processing. We release TabGenie as a PyPI package and provide its open-source code and a live demo at https://github.com/kasnerz/tabgenie.
Submitted to ACL 2023 System Demonstration Track
References in corpus (6)
- Training language models to follow instructions with human feedback
- Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning
- Multitask Prompted Training Enables Zero-Shot Task Generalization
- Learning to Reason for Text Generation from Scientific Tables
- MVP: Multi-task Supervised Pre-training for Natural Language Generation
- Innovations in Neural Data-to-text Generation: A Survey