3 papers
cs.LG2026
RIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention
Anderson R. Santos
Full self-attention in large language models scales as O(N^2), which limits long-context document analysis to 65,536 tokens and requires costly GPU clusters. The Reduced Interactio…
cs.LG2026
WISE-FM:Operation-Aware, Engineering-Informed Foundation Model for Multi-Task Well Design
Carine de Menezes Rebello, Anderson Rapello dos Santos, Idelfonso B. R. Nogueira
Deploying machine learning models across diverse well portfolios requires generalisation to wells with design parameters outside the training distribution. Current data-driven appr…
physics.chem-ph2025
ExPUFFIN: Thermodynamic Consistent Viscosity Prediction in an Extended Path-Unifying Feed-Forward Interfaced Network
Carine Menezes Rebello, Ulderico Di Caprio, Jenny Steen-Hansen +6
Accurate prediction of liquid viscosity is essential for process design and simulation, yet remains challenging for novel molecules. Conventional group-contribution models struggle…