◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andreas Chouliaras

University College Dublin

2 papers hereh-index 210 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.HC1
affiliations
  • University College Dublin

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.AI2026

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

Andreas Chouliaras, Luke Connolly, Dimitris Chatzpoulos

Training safe Reinforcement Learning (RL) systems is inherently challenging, with no guarantee of avoiding unwanted behaviors. The most effective defenses against this are (i) tran…

cs.HC2025

Maximizing the efficiency of human feedback in AI alignment: a comparative analysis

Andreas Chouliaras, Dimitris Chatzopoulos

Reinforcement Learning from Human Feedback (RLHF) relies on preference modeling to align machine learning systems with human values, yet the popular approach of random pair samplin…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.