6 papers
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
Yao Lu, Yuqi Li, Wenbin Xie +4
Although large language models (LLMs) have achieved revolutionary breakthroughs in many fields, their large model size and high computational cost pose significant challenges for p…
Few-shot Molecular Property Prediction: A Survey
Zeyu Wang, Tianyi Jiang, Huanchang Ma +6
AI-assisted molecular property prediction has become a promising technique in early-stage drug discovery and materials design in recent years. However, due to high-cost and complex…
DSPC: Dual-Stage Progressive Compression Framework for Efficient Long-Context Reasoning
Yaxin Gao, Yao Lu, Zongfei Zhang +3
Large language models (LLMs) have achieved remarkable success in many natural language processing (NLP) tasks. To achieve more accurate output, the prompts used to drive LLMs have…
LoRALib: A Standardized Benchmark for Evaluating LoRA-MoE Methods
Shaoheng Wang, Yao Lu, Yuqi Li +5
As a parameter efficient fine-tuning (PEFT) method, low-rank adaptation (LoRA) can save significant costs in storage and computing, but its strong adaptability to a single task is…
From LLM-anation to LLM-orchestrator: Coordinating Small Models for Data Labeling
Yao Lu, Zhaiyuan Ji, Jiawei Du +3
Although the annotation paradigm based on Large Language Models (LLMs) has made significant breakthroughs in recent years, its actual deployment still has two core bottlenecks: fir…
ReStNet: A Reusable & Stitchable Network for Dynamic Adaptation on IoT Devices
Maoyu Wang, Yao Lu, Jiaqi Nie +4
With the rapid development of deep learning, a growing number of pre-trained models have been publicly available. However, deploying these fixed models in real-world IoT applicatio…