◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jeffrey Flanigan

3 papers hereh-index 11 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.AI1
same name
  • Jeffrey Flanigan — 7 papers, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.AI2026

OSGuard: A Benchmark for Safety in Computer-Use Agents

Mina Mohammadmirzaei, Jeffrey Flanigan

Computer-use agents are increasingly evaluated by whether they complete realistic desktop and web tasks. However, task success alone can miss failures in which an agent reaches the…

cs.CL2026

Right or Wrong, Models Comply: Directional Blindness in LLM Moral Judgment

Jihye Kim, Jeffrey Flanigan

As language models take integrated roles across many domains, the response of LLMs to user pushback becomes a critical alignment property. Yet many existing evaluations treat compl…

cs.CL2026

Dialogue SWE-Bench: A Benchmark for Dialogue-Driven Coding Agents

Brendan King, Jeffrey Flanigan

AI coding agents have rapidly transformed software engineering, powering widely used interactive coding assistants. Despite their interactive real-world use, existing benchmarks ev…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.