1 citations
- Heidelberg UniversityDE1 paper
- LMU KlinikumDE1 paper
- Ludwig-Maximilians-Universität MünchenDE1 paper
- Machine ScienceUS1 paper
- Offenburg University of Applied SciencesDE1 paper
- Technische Hochschule MannheimDE1 paper
- University Hospital HeidelbergDE1 paper
- University Medical Centre MannheimDE1 paper
- University of TrentoIT1 paper
- University of VeronaIT1 paper
6 papers
Benchmarking Document Parsers on Mathematical Formula Extraction from PDFs
Pius Horn, Janis Keuper
Correctly parsing mathematical formulas from PDFs is critical for training large language models and building scientific knowledge bases from academic literature, yet existing benc…
WebMall -- A Multi-Shop Benchmark for Evaluating Web Agents
Ralph Peeters, Aaron Steiner, Luca Schwarz +2
LLM-based web agents have the potential to automate long-running web tasks, such as searching for products in multiple e-shops and subsequently ordering the cheapest products that…
Characterization of a novel plastic scintillation detector for in vivo electron dosimetry
Cornelius J. Bauer, Frank Schneider, Ida D. Göbel +3
Introduction: Real-time dosimetry of surface doses in electron beams has not been widely established yet. Plastic scintillation detectors (PSD) promise high spatial resolution and…
Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models
Georg Ahnert, Anna-Carolina Haensch, Barbara Plank +1
Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LLMs are typically trained to gene…
Treating Run-time Execution History as a First-Class Citizen: Co-Versioning Run-time Behavior alongside Code
Marcus Kessel
Behavioral Co-Versioning remains absent from mainstream practice: while developers routinely version source code with Git, they rarely persist and query how run-time behavior evolv…
Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs
Marco Bombieri, Simone Paolo Ponzetto, Marco Rospocher
Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects personas and demographic gro…