activity
20122024
most citedIterated risk measures for risk-sensitive Markov decision processes with discounted cost

4 citations · 9 across the 7 of their papers we have counts for

collaborators

8 papers

cs.LG2026

Distorted Distributional Policy Evaluation for Offline Reinforcement Learning

Ryo Iwaki, Takayuki Osogami

While Distributional Reinforcement Learning (DRL) methods have demonstrated strong performance in online settings, its success in offline scenarios remains limited. We hypothesize…

cs.GT2024

Socially efficient mechanism on the minimum budget

Hirota Kinoshita, Takayuki Osogami, Kohei Miyaguchi

In social decision-making among strategic agents, a universal focus lies on the balance between social and individual interests. Socially efficient mechanisms are thus desirably de…

cs.GT2023

Who Benefits from a Multi-Cloud Market? A Trading Networks Based Analysis

Segev Wasserkrug, Takayuki Osogami

In enterprise cloud computing, there is a big and increasing investment to move to multi-cloud computing, which allows enterprises to seamlessly utilize IT resources from multiple…

cs.LG2023

Regression with Sensor Data Containing Incomplete Observations

Takayuki Katsuki, Takayuki Osogami

This paper addresses a regression problem in which output label values are the results of sensing the magnitude of a phenomenon. A low value of such labels can mean either that the…

cs.GT20221 cited

Mechanism Learning for Trading Networks

Takayuki Osogami, Segev Wasserkrug, Elisheva S. Shamash

We study the problem of designing mechanisms for trading networks that satisfy four desired properties: dominant-strategy incentive compatibility, efficiency, weak budget balance (…

cs.GT20221 cited

Online Learning in Supply-Chain Games

Nicolò Cesa-Bianchi, Tommaso Cesari, Takayuki Osogami +2

We study a repeated game between a supplier and a retailer who want to maximize their respective profits without full knowledge of the problem parameters. After characterizing the…