5 papers · 1 filter
Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation
Katelyn Xiaoying Mei, Yi-Li Hsu, Minjoon Choi +5
Human evaluation plays a critical role in assessing the quality of generated text. However, the reliability and reproducibility of these evaluations depend on transparent and well-…
Leveraging Hierarchical Organization for Medical Multi-document Summarization
Yi-Li Hsu, Katelyn X. Mei, Lucy Lu Wang
Medical multi-document summarization (MDS) is a complex task that requires effectively managing cross-document relationships. This paper investigates whether incorporating hierarch…
Do Large Multimodal Models Solve Caption Generation for Scientific Figures? Lessons Learned from SciCap Challenge 2023
Ting-Yao E. Hsu, Yi-Li Hsu, Shaurya Rohatgi +8
Since the SciCap datasets launch in 2021, the research community has made significant progress in generating captions for scientific figures in scholarly articles. In 2023, the fir…
Is Explanation the Cure? Misinformation Mitigation in the Short Term and Long Term
Yi-Li Hsu, Shih-Chieh Dai, Aiping Xiong +1
With advancements in natural language processing (NLP) models, automatic explanation generation has been proposed to mitigate misinformation on social media platforms in addition t…
Label-Aware Hyperbolic Embeddings for Fine-grained Emotion Classification
Chih-Yao Chen, Tun-Min Hung, Yi-Li Hsu +1
Fine-grained emotion classification (FEC) is a challenging task. Specifically, FEC needs to handle subtle nuance between labels, which can be complex and confusing. Most existing m…