◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Nilay Pochhi

2 papers hereh-index 2155 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author1

Across the 1 of 2 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.CL1

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.AI2026

From RLHF to Direct Alignment: A Theoretical Unification of Preference Learning for Large Language Models

Tarun Raheja, Nilay Pochhi

Aligning large language models (LLMs) with human preferences has become essential for safe and beneficial AI deployment. While Reinforcement Learning from Human Feedback (RLHF) est…

cs.CL2024

Recent advancements in LLM Red-Teaming: Techniques, Defenses, and Ethical Considerations

Tarun Raheja, Nilay Pochhi, F. D. C. M. Curie

Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language processing tasks, but their vulnerability to jailbreak attacks poses significant security…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.