5 citations · 9 across the 9 of their papers we have counts for
4 papers · 1 filter
Patton: Language Model Pretraining on Text-Rich Networks
Bowen Jin, Wentao Zhang, Yu Zhang +4
A real-world text corpus sometimes comprises not only text documents but also semantic links between them (e.g., academic papers in a bibliographic network are linked by citations…
ReGen: Zero-Shot Text Classification via Training Data Generation with Progressive Dense Retrieval
Yue Yu, Yuchen Zhuang, Rongzhi Zhang +3
With the development of large language models (LLMs), zero-shot learning has attracted much attention for various NLP tasks. Different from prior works that generate training data…
Effective Seed-Guided Topic Discovery by Integrating Multiple Types of Contexts
Yu Zhang, Yunyi Zhang, Martin Michalski +3
Instead of mining coherent topics from a given text corpus in a completely unsupervised manner, seed-guided topic discovery methods leverage user-provided seed words to extract dis…
Few-Shot Fine-Grained Entity Typing with Automatic Label Interpretation and Instance Generation
Jiaxin Huang, Yu Meng, Jiawei Han
We study the problem of few-shot Fine-grained Entity Typing (FET), where only a few annotated entity mentions with contexts are given for each entity type. Recently, prompt-based t…