4 papers
Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction
Yujia Tong, Yuxi Wang, Yunyang Wan +3
Model compression techniques such as quantization and pruning are widely used to reduce the deployment cost of large language models (LLMs), with existing evaluations focusing almo…
A gentle tutorial on Bock's algorithm for minimum directed spanning trees with a structured reformulation
Yuxi Wang, Jungyeul Park
Bock's 1971 algorithm is an exact primal--dual method for the minimum-cost arborescence problem, but its Algol presentation obscures the interaction of its maintained arrays and la…
Advancing Text Classification with Large Language Models and Neural Attention Mechanisms
Ning Lyu, Yuxi Wang, Feng Chen +1
This study proposes a text classification algorithm based on large language models, aiming to address the limitations of traditional methods in capturing long-range dependencies, u…
Knowledge-Augmented Large Language Model Agents for Explainable Financial Decision-Making
Qingyuan Zhang, Yuxi Wang, Cancan Hua +2
This study investigates an explainable reasoning method for financial decision-making based on knowledge-enhanced large language model agents. To address the limitations of traditi…