reinforcement learning 2building energy management 1hvac control 1large language models 1model predictive control 1multi-zone vav 1reasoning models 1rollout verification 1TD3 1thermal energy storage 1
From the 2 of 2 linked papers with an AI index.
2 papers
cs.LG2026
Exact Action Values Are Not Enough: Rollout-Verified Reinforcement Fine-Tuning of a Reasoning Model for Multi-Zone VAV Control
Takumi Shioda, Kohei Terashima, Tatsuo Nagai
The paper investigates using a reasoning language model and TD3-guided reinforcement fine‑tuning to control multi‑zone variable‑air‑volume HVAC systems without building‑specific tr…
cs.LG2026
Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control
Takumi Shioda, Kohei Terashima, Tatsuo Nagai
The paper fine‑tunes an open‑weight reasoning language model via reinforcement learning with verifiable rewards to schedule thermal energy storage in buildings, achieving emissions…