4 papers · 1 filter
From Parameters to Data: A Task-Parameter-Guided Fine-Tuning Pipeline for Efficient LLM Alignment
Hao Chen, Qi Zhang, Liyao Li +7
Adapting Large Language Models (LLMs) to specialized domains typically incurs high data and computational overhead. While prior efficiency efforts have largely treated data selecti…
KMLP: A Scalable Hybrid Architecture for Web-Scale Tabular Data Modeling
Mingming Zhang, Pengfei Shi, Zhiqing Xiao +8
Predictive modeling on web-scale tabular data with billions of instances and hundreds of heterogeneous numerical features faces significant scalability challenges. These features e…
Beyond Tree Models: A Hybrid Model of KAN and gMLP for Large-Scale Financial Tabular Data
Mingming Zhang, Jiahao Hu, Pengfei Shi +8
Tabular data plays a critical role in real-world financial scenarios. Traditionally, tree models have dominated in handling tabular data. However, financial datasets in the industr…
From Laws to Motivation: Guiding Exploration through Law-Based Reasoning and Rewards
Ziyu Chen, Zhiqing Xiao, Xinbei Jiang +1
Large Language Models (LLMs) and Reinforcement Learning (RL) are two powerful approaches for building autonomous agents. However, due to limited understanding of the game environme…