From the 1 of 8 linked papers with an AI index.
8 papers
EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge
Chengxi Liao, Tao Xu, Zulong Chen +7
The paper presents EntSQL, a benchmark designed to evaluate text-to-SQL models on enterprise tasks that require grounding in long, proprietary business documents, using a bilingual…
DeInfer: Efficient Parallel Inferencing for Decomposed Large Language Models
You-Liang Huang, Xinhao Huang, Chengxi Liao +1
Existing works on large language model (LLM) decomposition mainly focus on improving performance on downstream tasks, but they ignore the poor parallel inference performance when t…
Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance
Jihang Li, Qing Liu, Zulong Chen +4
Tax code prediction is a crucial yet underexplored task in automating invoicing and compliance management for large-scale e-commerce platforms. Each product must be accurately mapp…
Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation
Minping Chen, Bing Xu, Zulong Chen +4
Person-Job Fit (PJF) is a critical component for online recruitment. Existing approaches face several challenges, particularly in handling low-quality job descriptions and similar…
SoLA: Leveraging Soft Activation Sparsity and Low-Rank Decomposition for Large Language Model Compression
Xinhao Huang, You-Liang Huang, Zeyi Wen
Large language models (LLMs) have demonstrated impressive capabilities across various tasks, but the billion-scale parameters pose deployment challenges. Although existing methods…
DOne: Decoupling Structure and Rendering for High-Fidelity Design-to-Code Generation
Xinhao Huang, Jinke Yu, Wenhao Xu +5
While Vision Language Models (VLMs) have shown promise in Design-to-Code generation, they suffer from a "holistic bottleneck-failing to reconcile high-level structural hierarchy wi…