3 papers
cs.LG2026
Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework
Zineb Senane, Axel Karlsson, Lele Cao +6
Existing evaluations of tabular synthesis models rely primarily on low-order statistics and downstream task performance, leaving multivariate causal relationships that go beyond pa…
cs.LG2025
Frequency Matters: When Time Series Foundation Models Fail Under Spectral Shift
Tianze Wang, Sofiane Ennadir, John Pertoft +7
Time series foundation models (TSFMs) have shown strong results on public benchmarks, prompting comparisons to a "BERT moment" for time series. Their effectiveness in industrial se…
cs.CL2025
GenCeption: Evaluate Vision LLMs with Unlabeled Unimodal Data
Lele Cao, Valentin Buchner, Zineb Senane +1
Multimodal Large Language Models (MLLMs) are typically assessed using expensive annotated multimodal benchmarks, which often lag behind the rapidly evolving demands of MLLM evaluat…