4 papers
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems
Stefano Cirillo, Domenico Desiato, Giuseppe Polese +1
We benchmark Google Embeddings (GE2), a Vertex-AI-hosted bi-encoder with 2,048-token context and explicit task-type conditioning, against five open-source alternatives: BGE-M3, E5-…
Bug Detective and Quality Coach: Developers' Mental Models of AI-Assisted IDE Tools
Paolo Buono, Mary Cerullo, Stefano Cirillo +7
AI-assisted tools support developers in performing cognitively demanding tasks such as bug detection and code readability assessment. Despite the advancements in the technical char…
Password Strength Analysis Through Social Network Data Exposure: A Combined Approach Relying on Data Reconstruction and Generative Models
Maurizio Atzori, Eleonora Calò, Loredana Caruccio +3
Although passwords remain the primary defense against unauthorized access, users often tend to use passwords that are easy to remember. This behavior significantly increases securi…
Augmenting Anonymized Data with AI: Exploring the Feasibility and Limitations of Large Language Models in Data Enrichment
Stefano Cirillo, Domenico Desiato, Giuseppe Polese +2
Large Language Models (LLMs) have demonstrated advanced capabilities in both text generation and comprehension, and their application to data archives might facilitate the privatiz…