3 papers
cs.LG2026
DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation
Peyman Hosseini, Ondrej Bohdal, Ahmed Alajrami +6
Large Language Model (LLM)-based agents can solve complex procedural tasks by interacting with environments over multiple turns, but this ability typically depends on large models,…
cs.CL2025
Fine-Tuning on Noisy Instructions: Effects on Generalization and Performance
Ahmed Alajrami, Xingwei Tan, Nikolaos Aletras
Instruction-tuning plays a vital role in enhancing the task-solving abilities of large language models (LLMs), improving their usability in generating helpful responses on various…
cs.CL2022
How does the pre-training objective affect what large language models learn about linguistic properties?
Ahmed Alajrami, Nikolaos Aletras
Several pre-training objectives, such as masked language modeling (MLM), have been proposed to pre-train language models (e.g. BERT) with the aim of learning better language repres…