2 papers
cs.LG2026
EntmaxKV: Support-Aware Decoding for Entmax Attention
Gonçalo Duarte, Miguel Couceiro, Marcos V. Treviso
Long-context decoding is increasingly limited by KV-cache memory traffic since each generated token attends over a cache whose size grows linearly with context length. Existing spa…
cs.AI2026
Which Are the Low-Resource Languages of the Semantic Web?
Ndeye-Emilie Mbengue, Pierre Monnin, Miguel Couceiro +1
Emerging digital technologies are exacerbating the existing divide in Open Access Data (OAD) between high-and low-resource languages, excluding many communities from the global dig…