◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Dmitrii Volkov

6 papers hereh-index 212 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author2
  • last author3

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.PL2
  • cs.AI1
  • cs.CR1
same name
  • Dmitrii Volkov — 6 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2025

Resurrecting saturated LLM benchmarks with adversarial encoding

Igor Ivanov, Dmitrii Volkov

Recent work showed that small changes in benchmark questions can reduce LLMs' reasoning and recall. We explore two such changes: pairing questions and adding more answer options, o…

cs.LG2024

Badllama 3: removing safety finetuning from Llama 3 in minutes

Dmitrii Volkov

We show that extensive LLM safety fine-tuning is easily subverted when an attacker has access to model weights. We evaluate three state-of-the-art fine-tuning methods-QLoRA, ReFT,…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.