4 papers
What to Format and How: A Benchmark and Workflow Approach for Document Formatting
Shihao Rao, Liang Li, Jiapeng Liu +6
Recent advances in large language models (LLMs) have opened up new possibilities for automated document formatting. However, real-world formatting often requires identifying target…
EXaMCaP: Subset Selection with Entropy Gain Maximization for Probing Capability Gains of Large Chart Understanding Training Sets
Jiapeng Liu, Liang Li, Bing Li +5
Recent works focus on synthesizing Chart Understanding (ChartU) training sets to inject advanced chart knowledge into Multimodal Large Language Models (MLLMs), where the sufficienc…
Reconstruction of Differentially Private Text Sanitization via Large Language Models
Shuchao Pang, Zhigang Lu, Haichen Wang +3
Differential privacy (DP) is the de facto privacy standard against privacy leakage attacks, including many recently discovered ones against large language models (LLMs). However, w…
Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion
Wei Dai, Peng Fu, Chunjing Gan
In an era marked by robust technological growth and swift information renewal, furnishing researchers and the populace with top-tier, avant-garde academic insights spanning various…