3 papers
cs.AI2026
ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability
Juarez Monteiro, Nathan Gavenski, Guilherme Lima +3
Reinforcement learning agents operating under partial observability must act on incomplete information, making them natural candidates for guidance from small language models (SLMs…
cs.AI2026
When in Doubt, Plan It Out: Committed Small Language Model Deliberation for Reactive Reinforcement Learning
Nathan Gavenski, Juarez Monteiro, Francisco Galuppo +2
Reinforcement Learning (RL) policies often degrade in unfamiliar environments because they lack explicit deliberation. We propose Plan, Align, Commit, Think (PACT), a hybrid archit…
cs.SI2024
I've Heard This Before: Initial Results on Tiktok's Impact On the Re-Popularization of Songs
Breno Matos, Francisco Galuppo, Rennan Cordeiro +1
With over a billion active users, TikTok's video-sharing service is currently one of the largest social media websites. This rise in TikTok's popularity has made the website a cent…