4 papers
Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts
Farhad Nooralahzadeh, Omid Rohanian, Yi Zhang +2
When a Vision-Language Model (VLM) sees a blue banana and answers "yellow", is the problem of perception or arbitration? We explore the question in ten VLMs with various sizes and…
Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures
Mauricio Fadel Argerich, Jonathan Fürst, Marta Patiño-Martínez
While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient LLM inference deployments tha…
Bench360: Benchmarking Local LLM Inference from 360 Degrees
Linus Stuhlmann, Mauricio Fadel Argerich, Jonathan Fürst
Running LLMs locally has become increasingly common, but users face a complex design space across models, quantization levels, inference engines, and serving scenarios. Existing in…
AgenticIE: An Adaptive Agent for Information Extraction from Complex Regulatory Documents
Gaye Colakoglu, Gürkan Solmaz, Jonathan Fürst
Declaration of Performance (DoP) documents, mandated by EU regulation, specify characteristics of construction products, such as fire resistance and insulation. While this informat…