3 citations · 4 across the 3 of their papers we have counts for
3 papers
Attention Lens: A Tool for Mechanistically Interpreting the Attention Head Information Retrieval Mechanism
Mansi Sakarvadia, Arham Khan, Aswathy Ajith +5
Transformer-based Large Language Models (LLMs) are the state-of-the-art for natural language tasks. Recent work has attempted to decode, by reverse engineering the role of linear l…
Adversarial Predictions of Data Distributions Across Federated Internet-of-Things Devices
Samir Rajani, Dario Dematties, Nathaniel Hudson +4
Federated learning (FL) is increasingly becoming the default approach for training machine learning models across decentralized Internet-of-Things (IoT) devices. A key advantage of…
Hierarchical and Decentralised Federated Learning
Omer Rana, Theodoros Spyridopoulos, Nathaniel Hudson +4
Federated learning has shown enormous promise as a way of training ML models in distributed environments while reducing communication costs and protecting data privacy. However, th…