collaborators

6 papers

q-fin.GN2025

FedSight AI: Multi-Agent System Architecture for Federal Funds Target Rate Prediction

Yuhan Hou, Tianji Rao, Jeremy Tan +7

The Federal Open Market Committee (FOMC) sets the federal funds rate, shaping monetary policy and the broader economy. We introduce \emph{FedSight AI}, a multi-agent framework that…

cs.LG2025

PREMAP: A Unifying PREiMage APproximation Framework for Neural Networks

Xiyue Zhang, Benjie Wang, Marta Kwiatkowska +1

Most methods for neural network verification focus on bounding the image, i.e., set of outputs for a given input set. This can be used to, for example, check the robustness of neur…

cs.CL2025

SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines

P Team, Xinrun Du, Yifan Yao +94

Large language models (LLMs) have demonstrated remarkable proficiency in mainstream academic disciplines such as mathematics, physics, and computer science. However, human knowledg…

cs.LG2024

Risk-Averse Certification of Bayesian Neural Networks

Xiyue Zhang, Zifan Wang, Yulong Gao +3

In light of the inherently complex and dynamic nature of real-world environments, incorporating risk measures is crucial for the robustness evaluation of deep learning models. In t…

cs.SE2024

FAST: Boosting Uncertainty-based Test Prioritization Methods for Neural Networks via Feature Selection

Jialuo Chen, Jingyi Wang, Xiyue Zhang +4

Due to the vast testing space, the increasing demand for effective and efficient testing of deep neural networks (DNNs) has led to the development of various DNN test case prioriti…

cs.LG2024

Automated Design of Linear Bounding Functions for Sigmoidal Nonlinearities in Neural Networks

Matthias König, Xiyue Zhang, Holger H. Hoos +2

The ubiquity of deep learning algorithms in various applications has amplified the need for assuring their robustness against small input perturbations such as those occurring in a…