3 papers
cs.LG2026
Neuron-Aware Data Selection for Annotation-Free LLM Self-Distillation
Zhuowei Chen, Xiang Lorraine Li
Post-training large language models (LLMs) without real-world interaction feedback or human-labeled supervision remains challenging, particularly in specialized domains where exper…
cs.LG2026
Neuron-Aware Active Few-Shot Learning for LLMs
Zhuowei Chen, Liwei Chen, Christian Schunn +2
Active Few-Shot Learning (AFSL) adapts LLMs to specialized domains by identifying the most valuable unlabeled samples for annotation and use as few-shot demonstrations, effectively…
cs.LG2026
Task-Relevant Representation Decoupling for Visual Reinforcement Learning Generalization
Jinwen Wang, Youfang Lin, Xiaobo Hu +4
Visual Reinforcement Learning (VRL) has achieved considerable success in solving control tasks. However, generalizing learned policies to new environments remains a major challenge…