3 papers
cs.AI2026
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences
Bang Nguyen, Dominik Soós, Qian Ma +8
The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on the computational aspect of thi…
cs.DL2025
CC30k: A Citation Contexts Dataset for Reproducibility-Oriented Sentiment Analysis
Rochana R. Obadage, Sarah M. Rajtmajer, Jian Wu
Sentiments about the reproducibility of cited papers in downstream literature offer community perspectives and have shown as a promising signal of the actual reproducibility of pub…
cs.DL2025
Toward Robust URL Extraction for Open Science: A Study of arXiv File Formats and Temporal Trends
Rochana R. Obadage, Lamia Salsabil, Sawood Alam +4
In this work, we study how URL extraction results depend on input format. We compiled a pilot dataset by extracting URLs from 10 arXiv papers and used the same heuristic method to…