6 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CL2026★ 1 cited
Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications
Ziyi Tong, Feifei Sun, Le Minh Nguyen
Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data grow, concerns about Pretraining…
cs.CL2018★ 6 cited
Sentence Modeling via Multiple Word Embeddings and Multi-level Comparison for Semantic Textual Similarity
Huy Nguyen Tien, Minh Nguyen Le, Yamasaki Tomohiro +1
Different word embedding models capture different aspects of linguistic properties. This inspired us to propose a model (M-MaxLSTM-CNN) for employing multiple sets of word embeddin…