7 citations · 13 across the 7 of their papers we have counts for
7 papers
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task
Yuya Fujisaki, Shiro Takagi, Hideki Asoh +1
The progress in text summarization techniques has been remarkable. However the task of accurately extracting and summarizing necessary information from highly specialized documents…
Towards Autonomous Hypothesis Verification via Language Models with Minimal Guidance
Shiro Takagi, Ryutaro Yamauchi, Wataru Kumagai
Research automation efforts usually employ AI as a tool to automate specific tasks within the research process. To create an AI that truly conduct research themselves, it must inde…
LPML: LLM-Prompting Markup Language for Mathematical Reasoning
Ryutaro Yamauchi, Sho Sonoda, Akiyoshi Sannai +1
In utilizing large language models (LLMs) for mathematical reasoning, addressing the errors in the reasoning and calculation present in the generated text by LLMs is a crucial chal…
Regularization and Variance-Weighted Regression Achieves Minimax Optimality in Linear MDPs: Theory and Practice
Toshinori Kitamura, Tadashi Kozuno, Yunhao Tang +12
Mirror descent value iteration (MDVI), an abstraction of Kullback-Leibler (KL) and entropy-regularized reinforcement learning (RL), has served as the basis for recent high-performi…
Learning Bound for Parameter Transfer Learning
Wataru Kumagai
We consider a transfer-learning problem by using the parameter transfer approach, where a suitable parameter of feature mapping is learned through one task and applied to another o…
Parallel Distributed Block Coordinate Descent Methods based on Pairwise Comparison Oracle
Kota Matsui, Wataru Kumagai, Takafumi Kanamori
This paper provides a block coordinate descent algorithm to solve unconstrained optimization problems. In our algorithm, computation of function values or gradients is not required…