◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Chatzopoulos

2 papers hereh-index 334 citations16 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.HC1
same name
  • D. Chatzopoulos — 1 paper, h 7

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.AI2026

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

Andreas Chouliaras, Luke Connolly, Dimitris Chatzpoulos

Training safe Reinforcement Learning (RL) systems is inherently challenging, with no guarantee of avoiding unwanted behaviors. The most effective defenses against this are (i) tran…

cs.HC2025

Maximizing the efficiency of human feedback in AI alignment: a comparative analysis

Andreas Chouliaras, Dimitris Chatzopoulos

Reinforcement Learning from Human Feedback (RLHF) relies on preference modeling to align machine learning systems with human values, yet the popular approach of random pair samplin…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.