◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Pankayaraj Pathmanathan

9 papers hereh-index 374 citations14 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author7
  • middle author2

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • cs.LG7
  • cs.CL1
  • cs.IR1

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedIs poisoning a real threat to LLM alignment? Maybe more so than you think

3 citations · 6 across the 8 of their papers we have counts for

collaborators
Showing 2026 · cs.LGShow all

2 papers · 2 filters

cs.LG2026

Compliance2LoRA: Personalizable On-Demand Safety Alignment on Arbitrary Policy Subsets via Hypernetwork-Generated LoRA Adapters

Pankayaraj Pathmanathan, Furong Huang

Post-training alignment in large reasoning models (LRMs) has significantly improved their adaptability to diverse safety compliance settings. However, as LRMs personalization for d…

cs.LG2026

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

Pankayaraj Pathmanathan, Furong Huang

While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted shortcomings due to the shallo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.