3 papers
cs.LG2026
In Two Minds about Lifelong Learning: Exploring Hemispheric Redundancy and Specialisation in Neural Models
Benjamin Smith, Levin Kuhlmann, Kaushik Roy +1
Persistent intelligent systems require the ability to learn continually, but current machine learning approaches face significant challenges in this area compared to biological lea…
cs.LG2026
MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning
Ankita Awasthi, Marco Apolinario, Kaushik Roy
In Online Continual Learning (OCL), a neural network sequentially learns from a non-stationary data stream in a single-pass with access only to a limited memory replay buffer. This…
cs.LG2024
Eigen Attention: Attention in Low-Rank Space for KV Cache Compression
Utkarsh Saxena, Gobinda Saha, Sakshi Choudhary +1
Large language models (LLMs) represent a groundbreaking advancement in the domain of natural language processing due to their impressive reasoning abilities. Recently, there has be…