1 citations · 1 across the 7 of their papers we have counts for
3 papers · 1 filter
Benchmark for Assessing Olfactory Perception of Large Language Models
Eftychia Makri, Nikolaos Nakis, Laura Sisson +4
Here we introduce the Olfactory Perception (OP) benchmark, designed to assess the capability of large language models (LLMs) to reason about smell. The benchmark contains 1,010 que…
MTBench: A Multimodal Time Series Benchmark for Temporal Reasoning and Question Answering
Jialin Chen, Aosong Feng, Ziyu Zhao +7
Understanding the relationship between textual news and time-series evolution is a critical yet under-explored challenge in applied data science. While multimodal learning has gain…
Long Sequence Modeling with Attention Tensorization: From Sequence to Tensor Learning
Aosong Feng, Rex Ying, Leandros Tassiulas
As the demand for processing extended textual data grows, the ability to handle long-range dependencies and maintain computational efficiency is more critical than ever. One of the…