3 papers
cs.AI2026
HypoBench: Towards Systematic and Principled Benchmarking for Hypothesis Generation
Haokun Liu, Sicong Huang, Jingyu Hu +2
There is growing interest in hypothesis generation with large language models (LLMs). However, fundamental questions remain: what makes a good hypothesis, and how can we systematic…
cs.CL2025
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
Sicong Huang, Qianqi Yan, Shengze Wang +1
Abstractive summarization using large language models (LLMs) has become an essential tool for condensing information. However, despite their ability to generate fluent summaries, t…
cs.CL2025
UCSC at SemEval-2025 Task 3: Context, Models and Prompt Optimization for Automated Hallucination Detection in LLM Output
Sicong Huang, Jincheng He, Shiyuan Huang +3
Hallucinations pose a significant challenge for large language models when answering knowledge-intensive queries. As LLMs become more widely adopted, it is crucial not only to dete…