5 papers
JFinTEB: Japanese Financial Text Embedding Benchmark
Masahiro Suzuki, Hiroki Sakaji
We introduce JFinTEB, the first comprehensive benchmark specifically designed for evaluating Japanese financial text embeddings. Existing embedding benchmarks provide limited cover…
Economy Watchers Survey Provides Datasets and Tasks for Japanese Financial Domain
Masahiro Suzuki, Hiroki Sakaji
Natural language processing (NLP) tasks in English and general domains are widely available and are often used to evaluate pre-trained language models. In contrast, fewer tasks are…
Refined and Segmented Price Sentiment Indices from Survey Comments
Masahiro Suzuki, Hiroki Sakaji
We aim to enhance a price sentiment index and to more precisely understand price trends from the perspective of not only consumers but also businesses. We extract comments related…
Interactive DualChecker for Mitigating Hallucinations in Distilling Large Language Models
Meiyun Wang, Masahiro Suzuki, Hiroki Sakaji +1
Large Language Models (LLMs) have demonstrated exceptional capabilities across various machine learning (ML) tasks. Given the high costs of creating annotated datasets for supervis…
JaFIn: Japanese Financial Instruction Dataset
Kota Tanabe, Masahiro Suzuki, Hiroki Sakaji +1
We construct an instruction dataset for the large language model (LLM) in the Japanese finance domain. Domain adaptation of language models, including LLMs, is receiving more atten…