◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

J. Castleman

3 papers hereh-index 314 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CY2
  • cs.LG1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.CY2026

Measuring Validity in LLM-based Resume Screening

Jane Castleman, Zeyu Shen, Blossom Metevier +2

Resume screening is perceived as a particularly suitable task for LLMs given their ability to analyze natural language; thus many entities rely on general purpose LLMs without furt…

cs.LG2026

The Geometry of Alignment Collapse: When Fine-Tuning Breaks Safety

Max Springer, Chung Peng Lee, Blossom Metevier +5

Fine-tuning aligned language models on benign tasks unpredictably degrades safety guardrails, even when training data contains no harmful content and developers have no adversarial…

cs.CY2025

Adultification Bias in LLMs and Text-to-Image Models

Jane Castleman, Aleksandra Korolova

The rapid adoption of generative AI models in domains such as education, policing, and social media raises significant concerns about potential bias and safety issues, particularly…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.