3 papers
cs.LG2026
Learning Rate Scaling across LoRA Ranks and Transfer to Full Finetuning
Nan Chen, Soledad Villar, Soufiane Hayou
Low-Rank Adaptation (LoRA) is a standard tool for parameter-efficient finetuning of large models. While it induces a small memory footprint, its training dynamics can be surprising…
cs.CL2025
SynerGen: Contextualized Generative Recommender for Unified Search and Recommendation
Vianne R. Gao, Chen Xue, Marc Versage +11
The dominant retrieve-then-rank pipeline in large-scale recommender systems suffers from mis-calibration and engineering overhead due to its architectural split and differing optim…
cs.LG2025
Exploring Pseudo-Token Approaches in Transformer Neural Processes
Jose Lara-Rangel, Nanze Chen, Fengzhe Zhang
Neural Processes (NPs) have gained attention in meta-learning for their ability to quantify uncertainty, together with their rapid prediction and adaptability. However, traditional…