most citedPlaying Atari Games with Deep Reinforcement Learning and Human Checkpoint Replay

76 citations · 76 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CL2024

Improving Legal Judgement Prediction in Romanian with Long Text Encoders

Mihai Masala, Traian Rebedea, Horia Velicu

In recent years,the entire field of Natural Language Processing (NLP) has enjoyed amazing novel results achieving almost human-like performance on a variety of tasks. Legal NLP dom…

cs.CL20237 cited

NeMo Guardrails: A Toolkit for Controllable and Safe LLM Applications with Programmable Rails

Traian Rebedea, Razvan Dinu, Makesh Sreedhar +2

NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems. Guardrails (or rails for short) are a specific way of contr…

cs.AI2023

Explaining Vision and Language through Graphs of Events in Space and Time

Mihai Masala, Nicolae Cudlenco, Traian Rebedea +1

Artificial Intelligence makes great advances today and starts to bridge the gap between vision and language. However, we are still far from understanding, explaining and controllin…

cs.CL2023

UPB at IberLEF-2023 AuTexTification: Detection of Machine-Generated Text using Transformer Ensembles

Andrei-Alexandru Preda, Dumitru-Clementin Cercel, Traian Rebedea +1

This paper describes the solutions submitted by the UPB team to the AuTexTification shared task, featured as part of IberLEF-2023. Our team participated in the first subtask, ident…

cs.CL2023

GEST: the Graph of Events in Space and Time as a Common Representation between Vision and Language

Mihai Masala, Nicolae Cudlenco, Traian Rebedea +1

One of the essential human skills is the ability to seamlessly build an inner representation of the world. By exploiting this representation, humans are capable of easily finding c…

cs.AI201676 cited

Playing Atari Games with Deep Reinforcement Learning and Human Checkpoint Replay

Ionel-Alexandru Hosu, Traian Rebedea

This paper introduces a novel method for learning how to play the most difficult Atari 2600 games from the Arcade Learning Environment using deep reinforcement learning. The propos…