◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Takumi Shioda

2 papers hereh-index 00 citations2 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG2

identity via Semantic Scholar / OpenAlex

works on
reinforcement learning 2building energy management 1hvac control 1large language models 1model predictive control 1multi-zone vav 1reasoning models 1rollout verification 1TD3 1thermal energy storage 1

From the 2 of 2 linked papers with an AI index.

collaborators

2 papers

cs.LG2026

Exact Action Values Are Not Enough: Rollout-Verified Reinforcement Fine-Tuning of a Reasoning Model for Multi-Zone VAV Control

Takumi Shioda, Kohei Terashima, Tatsuo Nagai

The paper investigates using a reasoning language model and TD3-guided reinforcement fine‑tuning to control multi‑zone variable‑air‑volume HVAC systems without building‑specific tr…

cs.LG2026

Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control

Takumi Shioda, Kohei Terashima, Tatsuo Nagai

The paper fine‑tunes an open‑weight reasoning language model via reinforcement learning with verifiable rewards to schedule thermal energy storage in buildings, achieving emissions…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.