5 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CR2024
Ignore Me But Don't Replace Me: Utilizing Non-Linguistic Elements for Pretraining on the Cybersecurity Domain
Eugene Jang, Jian Cui, Dayeon Yim +4
Cybersecurity information is often technically complex and relayed through unstructured text, making automation of cyber threat intelligence highly challenging. For such text domai…
cs.CL2023★ 3 cited
DarkBERT: A Language Model for the Dark Side of the Internet
Youngjin Jin, Eugene Jang, Jian Cui +3
Recent research has suggested that there are clear differences in the language used in the Dark Web compared to that of the Surface Web. As studies on the Dark Web commonly require…
cs.CL2022★ 5 cited
Towards WinoQueer: Developing a Benchmark for Anti-Queer Bias in Large Language Models
Virginia K. Felkner, Ho-Chun Herbert Chang, Eugene Jang +1
This paper presents exploratory work on whether and to what extent biases against queer and trans people are encoded in large language models (LLMs) such as BERT. We also propose a…