7 citations · 16 across the 5 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2023
Advancing Beyond Identification: Multi-bit Watermark for Large Language Models
KiYoon Yoo, Wonhyuk Ahn, Nojun Kwak
We show the viability of tackling misuses of large language models beyond the identification of machine-generated text. While existing zero-bit watermark methods focus on detection…
cs.CL2023★ 5 cited
Robust Multi-bit Natural Language Watermarking through Invariant Features
KiYoon Yoo, Wonhyuk Ahn, Jiho Jang +1
Recent years have witnessed a proliferation of valuable original natural language contents found in subscription-based media outlets, web novel platforms, and outputs of large lang…
cs.CL2022★ 7 cited
Detection of Word Adversarial Examples in Text Classification: Benchmark and Baseline via Robust Density Estimation
KiYoon Yoo, Jangho Kim, Jiho Jang +1
Word-level adversarial attacks have shown success in NLP models, drastically decreasing the performance of transformer-based models in recent years. As a countermeasure, adversaria…