3 papers
cs.CV2025
Atlas: A Novel Pathology Foundation Model by Mayo Clinic, Charité, and Aignostics
Maximilian Alber, Stephan Tietz, Jonas Dippel +24
Recent advances in digital pathology have demonstrated the effectiveness of foundation models across diverse applications. In this report, we present Atlas, a novel vision foundati…
cs.LG2024
Simple and Scalable Strategies to Continually Pre-train Large Language Models
Adam Ibrahim, Benjamin Thérien, Kshitij Gupta +5
Large language models (LLMs) are routinely pre-trained on billions of tokens, only to start the process over again once new data becomes available. A much more efficient solution i…
cs.CL2024
Continual Learning Under Language Shift
Evangelia Gogoulou, Timothée Lesort, Magnus Boman +1
The recent increase in data and model scale for language model pre-training has led to huge training costs. In scenarios where new data become available over time, updating a model…