Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Safe Planning and Policy Optimization via World Model Learning
Artem Latyshev, Gregory Gorbov, Aleksandr I. Panov
Reinforcement Learning (RL) applications in real-world scenarios must prioritize safety and reliability, which impose strict constraints on agent behavior. Model-based RL leverages…
cs.AI2025
CrafText Benchmark: Advancing Instruction Following in Complex Multimodal Open-Ended World
Zoya Volovikova, Gregory Gorbov, Petr Kuderov +2
Following instructions in real-world conditions requires the ability to adapt to the world's volatility and entanglement: the environment is dynamic and unpredictable, instructions…