papers

Publications (32)

cs.CL2025

Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications

Jimin Huang, Mengxi Xiao, Dong Li +41

Financial LLMs hold promise for advancing financial tasks and domain-specific applications. However, they are limited by scarce corpora, weak multimodal capabilities, and narrow ev…

cs.AI2026

Herculean: An Agentic Benchmark for Financial Intelligence

Xueqing Peng, Zhuohan Xie, Yupeng Cao +60

As AI agents improve, the central question is no longer whether they can solve isolated well-defined financial tasks, but whether they can reliably carry out financial professional…

cs.CV2024

VidEgoThink: Assessing Egocentric Video Understanding Capabilities for Embodied AI

Sijie Cheng, Kechen Fang, Yangyang Yu +6

Recent advancements in Multi-modal Large Language Models (MLLMs) have opened new avenues for applications in Embodied AI. Building on previous work, EgoThink, we introduce VidEgoTh…

cs.CL2024

FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Making

Yangyang Yu, Zhiyuan Yao, Haohang Li +14

Large language models (LLMs) have demonstrated notable potential in conducting complex tasks and are increasingly utilized in various financial applications. However, high-quality…

cs.CL2026

Exploring how EFL students talk to and through AI to develop texts

David James Woo, Yangyang Yu, Yilin Huang +3

Generative Artificial Intelligence (AI) introduces new considerations for English as a foreign language (EFL) writing pedagogy. This study explores how students talk to and through…

cs.CL2025

MultiFinBen: Benchmarking Large Language Models for Multilingual and Multimodal Financial Application

Xueqing Peng, Lingfei Qian, Yan Wang +44

Real-world financial analysis involves information across multiple languages and modalities, from reports and news to scanned filings and meeting recordings. Yet most existing eval…

cs.CY2025

Exploring EFL Secondary Students' AI-generated Text Editing While Composition Writing

David James Woo, Yangyang Yu, Kai Guo

Generative Artificial Intelligence is transforming how English as a foreign language students write. Still, little is known about how students manipulate text generated by generati…

cs.AI2023

Actively learning a Bayesian matrix fusion model with deep side information

Yangyang Yu, Jordan W. Suchow

High-dimensional deep neural network representations of images and concepts can be aligned to predict human annotations of diverse stimuli. However, such alignment requires the cos…

cs.CE2024

INVESTORBENCH: A Benchmark for Financial Decision-Making Tasks with LLM-based Agent

Haohang Li, Yupeng Cao, Yangyang Yu +12

Recent advancements have underscored the potential of large language model (LLM)-based agents in financial decision-making. Despite this progress, the field currently encounters tw…

cs.CY2025

A vibe coding learning design to enhance EFL students' talking to, through, and about AI

David James Woo, Kai Guo, Yangyang Yu

This innovative practice article reports on the piloting of vibe coding (using natural language to create software applications with AI) for English as a Foreign Language (EFL) edu…

q-fin.PM2023

TradingGPT: Multi-Agent System with Layered Memory and Distinct Characters for Enhanced Financial Trading Performance

Yang Li, Yangyang Yu, Haohang Li +2

Large Language Models (LLMs), prominently highlighted by the recent evolution in the Generative Pre-trained Transformers (GPT) series, have displayed significant prowess across var…

physics.plasm-ph2026

Formulation and verification of multiscale gyrokinetic simulation of kinetic-MHD processes in toroidal plasmas

Xishuo Wei, Pengfei Liu, Gyungjin Choi +11

A comprehensive gyrokinetic simulation model has been implemented in the global toroidal gyrokinetic code (GTC) and verified for studying low-frequency waves and turbulence in magn…

q-fin.CP2023

FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Design

Yangyang Yu, Haohang Li, Zhi Chen +6

Recent advancements in Large Language Models (LLMs) have exhibited notable efficacy in question-answering (QA) tasks across diverse domains. Their prowess in integrating extensive…

cs.CE2026

AutoRedTrader: Autonomous Red Teaming of Trading Agents through Synthetic Misinformation Injection

Zhiwei Liu, Yangyang Yu, Yupeng Cao +8

LLM-based financial agents increasingly rely on both numerical market data and textual signals for sequential trading and stock prediction. However, financial misinformation often…

cs.CL2024

FinBen: A Holistic Financial Benchmark for Large Language Models

Qianqian Xie, Weiguang Han, Zhengyu Chen +31

LLMs have transformed NLP and shown promise in various fields, yet their potential in finance is underexplored due to a lack of comprehensive evaluation benchmarks, the rapid devel…

cs.AI2026

FinTrace: Holistic Trajectory-Level Evaluation of LLM Tool Calling for Long-Horizon Financial Tasks

Yupeng Cao, Haohang Li, Weijin Liu +11

Recent studies demonstrate that tool-calling capability enables large language models (LLMs) to interact with external environments for long-horizon financial tasks. While existing…

cs.CV2026

Adaptive Data Augmentation with Multi-armed Bandit: Sample-Efficient Embedding Calibration for Implicit Pattern Recognition

Minxue Tang, Yangyang Yu, Aolin Ding +3

Recognizing implicit visual and textual patterns is essential in many real-world applications of modern AI. However, tackling long-tail pattern recognition tasks remains challengin…

cs.CL2025

Truth Neurons

Haohang Li, Yupeng Cao, Yangyang Yu +2

Despite their remarkable success and deployment across diverse workflows, language models sometimes produce untruthful responses. Our limited understanding of how truthfulness is m…

nucl-th2025

Neural Network Construction of the Equation of State from Relativistic ab initio Calculations

Kangmin Chen, Xiaoying Qu, Hui Tong +2

Constraining the nuclear matter equation of state (EOS) beyond saturation density is a central goal of nuclear physics and astrophysics. While the relativistic Brueckner-Hartree-Fo…

cs.CL2025

Product vs. Process: Exploring EFL Students' Editing of AI-Generated Text for Expository Writing

David James Woo, Yangyang Yu, Kai Guo +2

Text generated by artificial intelligence (AI) chatbots is increasingly used in English as a foreign language (EFL) writing contexts, yet its impact on students' expository writing…

cs.CL2026

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions

Xianzhong Ding, Yangyang Yu, Changwei Liu +1

A frontier language model's acknowledged "helpful programming assistant" persona does not survive long agentic-coding sessions in the deployment regime that production products act…

cs.AI2025

FLAG-Trader: Fusion LLM-Agent with Gradient-based Reinforcement Learning for Financial Trading

Guojun Xiong, Zhiyang Deng, Keyi Wang +10

Large language models (LLMs) fine-tuned on multimodal financial data have demonstrated impressive reasoning capabilities in various financial tasks. However, they often struggle wi…

cs.DC2024

Kilometer-Level Coupled Modeling Using 40 Million Cores: An Eight-Year Journey of Model Development

Xiaohui Duan, Yuxuan Li, Zhao Liu +38

With current and future leading systems adopting heterogeneous architectures, adapting existing models for heterogeneous supercomputers is of urgent need for improving model resolu…

hep-ph2024

Single spin asymmetry in dihadron production in SIDIS

Ren Yang, Yangyang Yu, Qihang Zhou +3

The paper calculates the helicity-dependent dihadron fragmentation function (DiFF), by extending the dihadron spectator model and examine the single longitudinal spin asymmetry $A^…

nucl-th2024

Nuclear mass table in deformed relativistic Hartree-Bogoliubov theory in continuum, II: Even- nuclei

DRHBc Mass Table Collaboration, Peng Guo, Xiaojie Cao +80

The mass table in the deformed relativistic Hartree-Bogoliubov theory in continuum (DRHBc) with the PC-PK1 density functional has been established for even- nuclei with $8\le Z\…

cs.CL2026

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

Yupeng Cao, Chengyang He, Yangyang Yu +2

Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial progress in automated veracity assessment,…

q-bio.NC2022

Transition behavior of the seizure dynamics modulated by the astrocyte inositol triphosphate noise

JiaJia Li, Peihua Feng, Liang Zhao +5

Epilepsy is a neurological disorder with recurrent seizures of complexity and randomness. Until now, the mechanism of epileptic randomness has not been fully elucidated. Inspired b…

cs.CL2025

When Agents Trade: Live Multi-Market Trading Benchmark for LLM Agents

Lingfei Qian, Xueqing Peng, Yan Wang +14

Although Large Language Model (LLM)-based agents are increasingly used in financial trading, it remains unclear whether they can reason and adapt in live markets, as most studies t…

nucl-th2025

Benchmarking nuclear energy density functionals with new mass data

Xiaoying Qu, Kangmin Chen, Cong Pan +2

Nuclear masses play a crucial role in both nuclear physics and astrophysics, driving sustained efforts toward their precise experimental determination and reliable theoretical pred…

cs.CL2025

Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim Evidence Reasoning

Shashidhar Reddy Javaji, Yupeng Cao, Haohang Li +3

Large language models (LLMs) are increasingly being used for complex research tasks such as literature review, idea generation, and scientific paper analysis, yet their ability to…

q-fin.RM2025

RiskLabs: Predicting Financial Risk Using Large Language Model based on Multimodal and Multi-Sources Data

Yupeng Cao, Zhi Chen, Prashant Kumar +7

The integration of Artificial Intelligence (AI) techniques, particularly large language models (LLMs), in finance has garnered increasing academic attention. Despite progress, exis…

cs.CE2025

FinAudio: A Benchmark for Audio Large Language Models in Financial Applications

Yupeng Cao, Haohang Li, Yangyang Yu +10

Audio Large Language Models (AudioLLMs) have received widespread attention and have significantly improved performance on audio tasks such as conversation, audio understanding, and…