◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

A. Kalyan

9 papers hereh-index 144.5k citations36 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author8
  • last author1

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • cs.CL5
  • cs.LG3
  • cs.AI1

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2025

Engagement Undermines Safety: How Stereotypes and Toxicity Shape Humor in Language Models

Atharvan Dogra, Soumya Suvra Ghosal, Ameet Deshpande +2

Large language models are increasingly used for creative writing and engagement content, raising safety concerns about the outputs. Therefore, casting humor generation as a testbed…

cs.CL2025

Language Models can Subtly Deceive Without Lying: A Case Study on Strategic Phrasing in Legislation

Atharvan Dogra, Krishna Pillutla, Ameet Deshpande +5

We explore the ability of large language models (LLMs) to engage in subtle deception through strategically phrasing and intentionally manipulating information. This harmful behavio…

cs.CL2025

PersonaGym: Evaluating Persona Agents and LLMs

Vinay Samuel, Henry Peng Zou, Yue Zhou +6

Persona agents, which are LLM agents conditioned to act according to an assigned persona, enable contextually rich and user aligned interactions across domains like education and h…

cs.CL2025

Probing AI Safety with Source Code

Ujwal Narayan, Shreyas Chaudhari, Ashwin Kalyan +4

Large language models (LLMs) have become ubiquitous, interfacing with humans in numerous safety-critical applications. This necessitates improving capabilities, but importantly cou…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.