5 papers · 1 filter
Multi-domain Multi-modal Document Classification Benchmark with a Multi-level Taxonomy
Denghao Ma, Qing Liu, Zulong Chen +5
Document classification forms the backbone of modern enterprise content management, yet existing benchmarks remain trapped in oversimplified paradigms -- single domain settings wit…
Towards Privacy-Preserving Machine Translation at the Inference Stage: A New Task and Benchmark
Wei Shao, Lemao Liu, Yinqiao Li +3
Current online translation services require sending user text to cloud servers, posing a risk of privacy leakage when the text contains sensitive information. This risk hinders the…
DiffETM: Diffusion Process Enhanced Embedded Topic Model
Wei Shao, Mingyang Liu, Linqi Song
The embedded topic model (ETM) is a widely used approach that assumes the sampled document-topic distribution conforms to the logistic normal distribution for easier optimization.…
CLOMO: Counterfactual Logical Modification with Large Language Models
Yinya Huang, Ruixin Hong, Hongming Zhang +6
In this study, we delve into the realm of counterfactual reasoning capabilities of large language models (LLMs). Our primary objective is to cultivate the counterfactual thought pr…
Privacy in LLM-based Recommendation: Recent Advances and Future Directions
Sichun Luo, Wei Shao, Yuxuan Yao +9
Nowadays, large language models (LLMs) have been integrated with conventional recommendation models to improve recommendation performance. However, while most of the existing works…