13 papers
VetScore: Risk-Weighted Fact Verification for Veterinary Long-Form QA with Citations
Ivan KartáÄ, Ivan Kartáč, Jan Tovarys +3
Citation excerpts can be used to increase the reliability of generated outputs and their faithfulness to cited sources, which is especially important in high-stakes domains such as…
Strategies for Span Labeling with Large Language Models
Danil Semin, OndÅej DuÅ¡ek, ZdenÄk Kasner
Large language models (LLMs) are increasingly used for text analysis tasks, such as named entity recognition or error detection. Unlike encoder-based models, however, generative ar…
Can LLM Coding Agents Reason About Time Series?
Filip RechtorÃk, OndÅej DuÅ¡ek, ZdenÄk Kasner
Large language models (LLMs) are increasingly being used for automated decision-making systems in finance, healthcare, or environmental monitoring. Time series data are ubiquitous…
Modular Monolingual Adaptation using Pretrained Language Models
Nalin Kumar, OndÅej DuÅ¡ek
Building monolingual language models (LMs) for low-resource languages typically relies on adapting pretrained language models (PLMs) by finetuning the whole model on the target lan…
AnimatedLLM: Explaining LLMs with Interactive Visualizations
ZdenÄk Kasner, OndÅej DuÅ¡ek
Large language models (LLMs) are becoming central to natural language processing education, yet materials showing their mechanics are sparse. We present AnimatedLLM, an interactive…
LLM Agents Implement an NLG System from Scratch: Building Interpretable Rule-Based RDF-to-Text Generators
Mateusz Lango, OndÅej DuÅ¡ek
We present a novel neurosymbolic framework for RDF-to-text generation, in which the model is "trained" through collaborative interactions among multiple LLM agents rather than trad…