66 citations · 102 across the 27 of their papers we have counts for
Showing 2023 · cs.CLShow all
3 papers · 2 filters
cs.CL2023
Text Representation Distillation via Information Bottleneck Principle
Yanzhao Zhang, Dingkun Long, Zehan Li +1
Pre-trained language models (PLMs) have recently shown great success in text representation field. However, the high computational cost and high-dimensional representation of PLMs…
cs.CL2023
Language Models are Universal Embedders
Xin Zhang, Zehan Li, Yanzhao Zhang +4
In the large language model (LLM) revolution, embedding is a key component of various systems, such as retrieving knowledge or memories for LLMs or building content moderation filt…
cs.CL2023★ 66 cited
Towards General Text Embeddings with Multi-stage Contrastive Learning
Zehan Li, Xin Zhang, Yanzhao Zhang +3
We present GTE, a general-purpose text embedding model trained with multi-stage contrastive learning. In line with recent advancements in unifying various NLP tasks into a single f…