collaborators

6 papers

cs.LG2026

FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics

Qiran Zou, Hou Hei Lam, Wenhao Zhao +11

AI research agents accelerate ML research by automating hypothesis generation, experimentation, and empirical refinement. Existing agent strategies range from greedy hill-climbing…

cs.CL2026

HypoSpace: A Diagnostic Benchmark for Set-Valued Hypothesis Generation under Underdetermination and Sublinear Coverage Bounds

Tingting Chen, Beibei Lin, Zifeng Yuan +5

Many scientific problems are underdetermined: multiple distinct hypotheses are equally consistent with the same observations. In such settings, effective inference requires not onl…

cs.CL2026

FML-bench: Benchmarking Machine Learning Agents for Scientific Research

Qiran Zou, Hou Hei Lam, Wenhao Zhao +7

Large language models (LLMs) have sparked growing interest in machine learning research agents that can autonomously propose ideas and conduct experiments. However, existing benchm…

cs.CE2026

Navigating heterogeneous protein landscapes through geometry-aware smoothing

Srinivas Anumasa, Barath Chandran, Tingting Chen +15

The evolutionary fitness landscape of biological molecules is extremely sparse and heterogeneous, with functional sequences forming isolated dense ``islands'' within a vast combina…

cs.CY2026

AI-generated data contamination erodes pathological variability and diagnostic reliability

Hongyu He, Shaowen Xiang, Ye Zhang +15

Generative artificial intelligence (AI) is rapidly populating medical records with synthetic content, creating a feedback loop where future models are increasingly at risk of train…

cs.LG2025

Data-Dependent Smoothing for Protein Discovery with Walk-Jump Sampling

Srinivas Anumasa, Barath Chandran. C, Tingting Chen +1

Diffusion models have emerged as a powerful class of generative models by learning to iteratively reverse the noising process. Their ability to generate high-quality samples has ex…