works on

From the 1 of 9 linked papers with an AI index.

collaborators

9 papers

cs.AI2026

A Threshold Exceedance Framework for CBRN Uplift Evaluation in Frontier Language Models

Rahul Gupta, Abhinav Mohanty, Payal Motwani +8

The paper introduces a Threshold Exceedance Criteria (TEC) framework to systematically evaluate whether frontier language models increase a non‑expert's ability to plan chemical, b…

cs.CL2026

When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

Yuting Ning, Jaylen Jones, Zhehao Zhang +5

Computer-use agents (CUAs) have made tremendous progress in the past year, yet they still frequently produce misaligned actions that deviate from the user's original intent. Such m…

cs.CR2026

Evaluating Nova 2.0 Lite model under Amazon's Frontier Model Safety Framework

Satyapriya Krishna, Matteo Memelli, Tong Wang +5

Amazon published its Frontier Model Safety Framework (FMSF) as part of the Paris AI summit, following which we presented a report on Amazon's Premier model. In this report, we pres…

cs.AI2026

Embedded AI Companion System on Edge Devices

Rahul Gupta, Stephen D. H. Hsu

Computational resource constraints on edge devices make it difficult to develop a fully embedded AI companion system with a satisfactory user experience. AI companion and memory sy…

cs.CV2025

VMDT: Decoding the Trustworthiness of Video Foundation Models

Yujin Potter, Zhun Wang, Nicholas Crispino +11

As foundation models become more sophisticated, ensuring their trustworthiness becomes increasingly critical; yet, unlike text and image, the video modality still lacks comprehensi…

cs.CL2025

D-REX: A Benchmark for Detecting Deceptive Reasoning in Large Language Models

Satyapriya Krishna, Andy Zou, Rahul Gupta +6

The safety and alignment of Large Language Models (LLMs) are critical for their responsible deployment. Current evaluation methods predominantly focus on identifying and preventing…