6 citations · 6 across the 2 of their papers we have counts for
4 papers · 1 filter
Aggregate vs. Personalized Judges in Business Idea Evaluation: Evidence from Expert Disagreement
Wataru Hirota, Tomoki Taniguchi, Tomoko Ohkuma +6
Evaluating LLM-generated business ideas is often harder to scale than generating them. Unlike standard NLP benchmarks, business idea evaluation relies on multi-dimensional criteria…
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
Keisuke Ueda, Wataru Hirota, Takuto Asakura +4
Large language models (LLMs) are increasingly used to support creative tasks such as research idea generation. While recent work has shown that structured dialogues between LLMs ca…
HyperPIE: Hyperparameter Information Extraction from Scientific Publications
Tarek Saier, Mayumi Ohta, Takuto Asakura +1
Automatic extraction of information from publications is key to making scientific knowledge machine readable at a large scale. The extracted information can, for example, facilitat…
A Quantitative Evaluation of Natural Language Question Interpretation for Question Answering Systems
Takuto Asakura, Jin-Dong Kim, Yasunori Yamamoto +2
Systematic benchmark evaluation plays an important role in the process of improving technologies for Question Answering (QA) systems. While currently there are a number of existing…