4 citations · 4 across the 5 of their papers we have counts for
5 papers
Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations
Omar Benjelloun, Leonardo Martins Bianco, Isabelle Guyon +8
Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecified execution details and bri…
Stylized Meta-Album: Group-bias injection with style transfer to study robustness against distribution shifts
Romain Mussard, Aurélien Gauffre, Ihsan Ullah +4
We introduce Stylized Meta-Album (SMA), a new image classification meta-dataset comprising 24 datasets (12 content datasets, and 12 stylized datasets), designed to advance studies…
Usefulness of LLMs as an Author Checklist Assistant for Scientific Papers: NeurIPS'24 Experiment
Alexander Goldberg, Ihsan Ullah, Thanh Gia Hieu Khuong +4
Large language models (LLMs) represent a promising, but controversial, tool in aiding scientific peer review. This study evaluates the usefulness of LLMs in a conference setting as…
RelevAI-Reviewer: A Benchmark on AI Reviewers for Survey Paper Relevance
Paulo Henrique Couto, Quang Phuoc Ho, Nageeta Kumari +4
Recent advancements in Artificial Intelligence (AI), particularly the widespread adoption of Large Language Models (LLMs), have significantly enhanced text analysis capabilities. T…
Auto-survey Challenge
Thanh Gia Hieu Khuong, Benedictus Kent Rachmat
We present a novel platform for evaluating the capability of Large Language Models (LLMs) to autonomously compose and critique survey papers spanning a vast array of disciplines in…