2 papers
cs.AI2025
Who Judges the Judge? LLM Jury-on-Demand: Building Trustworthy LLM Evaluation Systems
Xiaochuan Li, Ke Wang, Girija Gouda +5
As Large Language Models (LLMs) become integrated into high-stakes domains, there is a growing need for evaluation methods that are both scalable for real-time deployment and relia…
cs.LG2024
Towards a framework on tabular synthetic data generation: a minimalist approach: theory, use cases, and limitations
Yueyang Shen, Agus Sudjianto, Arun Prakash R +5
We propose and study a minimalist approach towards synthetic tabular data generation. The model consists of a minimalistic unsupervised SparsePCA encoder (with contingent clusterin…