3 papers
cs.DB2026
Pneuma-Seeker: A Relational Reification Mechanism to Align AI Agents with Human Work over Relational Data
Muhammad Imam Luthfi Balaka, John Hillesland, Kemal Badur +1
When faced with data problems, many data workers cannot articulate their information need precisely enough for software to help. Although LLMs interpret natural-language requests,…
cs.CY2025
Emergent evaluation hubs in a decentralizing large language model ecosystem
Manuel Cebrian, Tomomi Kito, Raul Castro Fernandez
Large language models are proliferating, and so are the benchmarks that serve as their common yardsticks. We ask how the agglomeration patterns of these two layers compare: do they…
cs.CL2025
Mass-Scale Analysis of In-the-Wild Conversations Reveals Complexity Bounds on LLM Jailbreaking
Aldan Creo, Raul Castro Fernandez, Manuel Cebrian
As large language models (LLMs) become increasingly deployed, understanding the complexity and evolution of jailbreaking strategies is critical for AI safety. We present a mass-sca…