collaborators

10 papers

cs.CL2025

Efficient Post-Training Refinement of Latent Reasoning in Large Language Models

Xinyuan Wang, Dongjie Wang, Wangyang Ying +5

Reasoning is a key component of language understanding in Large Language Models. While Chain-of-Thought prompting enhances performance via explicit intermediate steps, it suffers f…

cs.LG2025

DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming

Arun Vignesh Malarkkan, Haoyue Bai, Anjali Kaushik +1

In real-world applications, domain data often contains identifiable or sensitive attributes, is subject to strict regulations (e.g., HIPAA, GDPR), and requires explicit data featur…

cs.LG2025

Data-Efficient Symbolic Regression via Foundation Model Distillation

Wangyang Ying, Jinghan Zhang, Haoyue Bai +5

Discovering interpretable mathematical equations from observed data (a.k.a. equation discovery or symbolic regression) is a cornerstone of scientific discovery, enabling transparen…

cs.LG2025

Where's the liability in the Generative Era? Recovery-based Black-Box Detection of AI-Generated Content

Haoyue Bai, Yiyou Sun, Wei Cheng +1

The recent proliferation of photorealistic images created by generative models has sparked both excitement and concern, as these images are increasingly indistinguishable from real…

cs.LG2025

Causal Graph Profiling via Structural Divergence for Robust Anomaly Detection in Cyber-Physical Systems

Arun Vignesh Malarkkan, Haoyue Bai, Dongjie Wang +1

With the growing complexity of cyberattacks targeting critical infrastructures such as water treatment networks, there is a pressing need for robust anomaly detection strategies th…

cs.LG2025

Incremental Causal Graph Learning for Online Cyberattack Detection in Cyber-Physical Infrastructures

Arun Vignesh Malarkkan, Dongjie Wang, Haoyue Bai +1

The escalating threat of cyberattacks on real-time critical infrastructures poses serious risks to public safety, demanding detection methods that effectively capture complex syste…