4 papers
MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models
Lionel Z. Wang, Ka Chung Ng, Yiming Ma +1
Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language models (LLMs), as part of generative AI,…
Modeling Inverse Ellipsometry Problem via Flow Matching with a Large-Scale Dataset
Yiming Ma, Jianzhi Teng, Xinjie Li +5
Inverse ellipsometry, i.e., reconstructing optical constants and film thickness from the measured phase difference and amplitude ratio , is a fundamentally ill-posed probl…
How Implicit Bias Accumulates and Propagates in LLM Long-term Memory
Yiming Ma, Lixu Wang, Lionel Z. Wang +6
Long-term memory mechanisms enable Large Language Models (LLMs) to maintain continuity and personalization across extended interaction lifecycles, but they also introduce new and u…
JiraiBench: A Bilingual Benchmark for Evaluating Large Language Models' Detection of Human Self-Destructive Behavior Content in Jirai Community
Yunze Xiao, Tingyu He, Lionel Z. Wang +6
This paper introduces JiraiBench, the first bilingual benchmark for evaluating large language models' effectiveness in detecting self-destructive content across Chinese and Japanes…