activity
20122024
most citedLearning Bound for Parameter Transfer Learning

7 citations · 13 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CL2024

A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task

Yuya Fujisaki, Shiro Takagi, Hideki Asoh +1

The progress in text summarization techniques has been remarkable. However the task of accurately extracting and summarizing necessary information from highly specialized documents…

cs.AI20231 cited

Towards Autonomous Hypothesis Verification via Language Models with Minimal Guidance

Shiro Takagi, Ryutaro Yamauchi, Wataru Kumagai

Research automation efforts usually employ AI as a tool to automate specific tasks within the research process. To create an AI that truly conduct research themselves, it must inde…

cs.AI20232 cited

LPML: LLM-Prompting Markup Language for Mathematical Reasoning

Ryutaro Yamauchi, Sho Sonoda, Akiyoshi Sannai +1

In utilizing large language models (LLMs) for mathematical reasoning, addressing the errors in the reasoning and calculation present in the generated text by LLMs is a crucial chal…

cs.LG2023

Regularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice

Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang +12

Mirror descent value iteration (MDVI), an abstraction of Kullback-Leibler (KL) and entropy-regularized reinforcement learning (RL), has served as the basis for recent high-performi…

stat.ML20167 cited

Learning Bound for Parameter Transfer Learning

Wataru Kumagai

We consider a transfer-learning problem by using the parameter transfer approach, where a suitable parameter of feature mapping is learned through one task and applied to another o…

stat.ML20141 cited

Parallel Distributed Block Coordinate Descent Methods based on Pairwise Comparison Oracle

Kota Matsui, Wataru Kumagai, Takafumi Kanamori

This paper provides a block coordinate descent algorithm to solve unconstrained optimization problems. In our algorithm, computation of function values or gradients is not required…