2 papers
cs.LG2026
Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem
Domagoj Ševerdija, Jurica Maltar, Nathan Chappel +1
Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly from raw instances. What such a…
cs.CL2023
Compressing Sentence Representation with maximum Coding Rate Reduction
Domagoj Ševerdija, Tomislav Prusina, Antonio Jovanović +3
In most natural language inference problems, sentence representation is needed for semantic retrieval tasks. In recent years, pre-trained large language models have been quite effe…