◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Balázs Varga

3 papers hereh-index 5112 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • math.OC1

identity via Semantic Scholar / OpenAlex

activity
20212023
most citedDeep Q-learning: a robust control approach

18 citations · 18 across the 3 of their papers we have counts for

collaborators

3 papers

math.OC2023

Controlled Descent Training

Viktor Andersson, Balázs Varga, Vincent Szolnoky +3

In this work, a novel and model-based artificial neural network (ANN) training method is developed supported by optimal control theory. The method augments training labels in order…

cs.LG2022★ 18 cited

Deep Q-learning: a robust control approach

Balazs Varga, Balazs Kulcsar, Morteza Haghir Chehreghani

In this paper, we place deep Q-learning into a control-oriented perspective and study its learning dynamics with well-established techniques from robust control. We formulate an un…

cs.LG2021

Constrained Policy Gradient Method for Safe and Fast Reinforcement Learning: a Neural Tangent Kernel Based Approach

Balázs Varga, Balázs Kulcsár, Morteza Haghir Chehreghani

This paper presents a constrained policy gradient algorithm. We introduce constraints for safe learning with the following steps. First, learning is slowed down (lazy learning) so…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.