Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
Prajit T Rajendran, Fabio Arnez, Huascar Espinoza +2
In high stakes environments, agents relying purely on imitation learning or reinforcement learning often struggle to avoid safety-critical errors during exploration. Existing reinf…
cs.LG2023
An Approach for Efficient Neural Architecture Search Space Definition
Léo Pouy, Fouad Khenfri, Patrick Leserf +2
As we advance in the fast-growing era of Machine Learning, various new and more complex neural architectures are arising to tackle problem more efficiently. On the one hand their e…