◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Davit Soselia

8 papers hereh-index 7271 citations14 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author5

Across the 8 of 8 papers where every author was matched, so the position is known.

fields
  • cs.CV5
  • cs.LG3

identity via Semantic Scholar / OpenAlex

activity
20182026
most citedReviving Shift Equivariance in Vision Transformers

1 citations · 3 across the 5 of their papers we have counts for

collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2024

OPTune: Efficient Online Preference Tuning

Lichang Chen, Jiuhai Chen, Chenxi Liu +6

Reinforcement learning with human feedback~(RLHF) is critical for aligning Large Language Models (LLMs) with human preference. Compared to the widely studied offline version of RLH…

cs.LG2024★ 1 cited

ODIN: Disentangled Reward Mitigates Hacking in RLHF

Lichang Chen, Chen Zhu, Davit Soselia +6

In this work, we study the issue of reward hacking on the response length, a challenge emerging in Reinforcement Learning from Human Feedback (RLHF) on LLMs. A well-formatted, verb…

cs.LG2018

Reproduction Report on "Learn to Pay Attention"

Levan Shugliashvili, Davit Soselia, Shota Amashukeli +1

We have successfully implemented the "Learn to Pay Attention" model of attention mechanism in convolutional neural networks, and have replicated the results of the original paper i…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.