activity
20192026
most citedA Robust Defense against Adversarial Attacks on Deep Learning-based Malware Detectors via (De)Randomized Smoothing

11 citations · 18 across the 18 of their papers we have counts for

collaborators
Showing cs.CLShow all

6 papers · 1 filter

cs.CL2026

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping

Yannis Belkhiter, Seshu Tirupathi, Giulio Zizzo +1

The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling LRMs to reason longer, and mo…

cs.CL2025

Step-Tagging: Toward controlling the generation of Language Reasoning Models through step monitoring

Yannis Belkhiter, Seshu Tirupathi, Giulio Zizzo +1

The field of Language Reasoning Models (LRMs) has been very active over the past few years with advances in training and inference techniques enabling LRMs to reason longer, and mo…

cs.CL2024

Granite Guardian

Inkit Padhi, Manish Nagireddy, Giandomenico Cornacchia +20

We introduce the Granite Guardian models, a suite of safeguards designed to provide risk detection for prompts and responses, enabling safe and responsible use in combination with…

cs.CL2024

HarmLevelBench: Evaluating Harm-Level Compliance and the Impact of Quantization on Model Alignment

Yannis Belkhiter, Giulio Zizzo, Sergio Maffeis

With the introduction of the transformers architecture, LLMs have revolutionized the NLP field with ever more powerful models. Nevertheless, their development came up with several…

cs.CL2024

Knowledge-Augmented Reasoning for EUAIA Compliance and Adversarial Robustness of LLMs

Tomas Bueno Momcilovic, Dian Balta, Beat Buesser +2

The EU AI Act (EUAIA) introduces requirements for AI systems which intersect with the processes required to establish adversarial robustness. However, given the ambiguous language…

cs.CL2023

Matching Pairs: Attributing Fine-Tuned Models to their Pre-Trained Large Language Models

Myles Foley, Ambrish Rawat, Taesung Lee +3

The wide applicability and adaptability of generative large language models (LLMs) has enabled their rapid adoption. While the pre-trained models can perform many tasks, such model…