10 papers
Efficient Post-Training Refinement of Latent Reasoning in Large Language Models
Xinyuan Wang, Dongjie Wang, Wangyang Ying +5
Reasoning is a key component of language understanding in Large Language Models. While Chain-of-Thought prompting enhances performance via explicit intermediate steps, it suffers f…
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
Arun Vignesh Malarkkan, Haoyue Bai, Anjali Kaushik +1
In real-world applications, domain data often contains identifiable or sensitive attributes, is subject to strict regulations (e.g., HIPAA, GDPR), and requires explicit data featur…
Data-Efficient Symbolic Regression via Foundation Model Distillation
Wangyang Ying, Jinghan Zhang, Haoyue Bai +5
Discovering interpretable mathematical equations from observed data (a.k.a. equation discovery or symbolic regression) is a cornerstone of scientific discovery, enabling transparen…
Where's the liability in the Generative Era? Recovery-based Black-Box Detection of AI-Generated Content
Haoyue Bai, Yiyou Sun, Wei Cheng +1
The recent proliferation of photorealistic images created by generative models has sparked both excitement and concern, as these images are increasingly indistinguishable from real…
Causal Graph Profiling via Structural Divergence for Robust Anomaly Detection in Cyber-Physical Systems
Arun Vignesh Malarkkan, Haoyue Bai, Dongjie Wang +1
With the growing complexity of cyberattacks targeting critical infrastructures such as water treatment networks, there is a pressing need for robust anomaly detection strategies th…
Incremental Causal Graph Learning for Online Cyberattack Detection in Cyber-Physical Infrastructures
Arun Vignesh Malarkkan, Dongjie Wang, Haoyue Bai +1
The escalating threat of cyberattacks on real-time critical infrastructures poses serious risks to public safety, demanding detection methods that effectively capture complex syste…