From the 1 of 9 linked papers with an AI index.
9 papers
Mechanistic Evidence for Preserved-but-Misaligned Representations in Non-IID FedAvg
Muhammad Haseeb, Salaar Masood, Muhammad Abdullah Sohail +2
The paper investigates why federated averaging (FedAvg) performs poorly on non‑IID client data, finding that client models retain useful internal representations but these become m…
DocAtlas: Multilingual Document Understanding Across 80+ Languages
Ahmed Heakl, Youssef Mohamed, Abdullah Sohail +6
Multilingual document understanding remains limited for low-resource languages due to scarce training data and model-based annotation pipelines that perpetuate existing biases. We…
Latent-DARM: Bridging Discrete Diffusion And Autoregressive Models For Reasoning
Lina Berrayana, Ahmed Heakl, Abdullah Sohail +3
Most multi-agent systems rely exclusively on autoregressive language models (ARMs) that are based on sequential generation. Although effective for fluent text, ARMs limit global re…
Towards Misinformation Resilience in Pakistan: A Participatory Study with Low-Socioeconomic Status Adults
Muhammad Abdullah Sohail, Amna Hassan, Shaheer Hammad +2
Digital misinformation disproportionately affects low-socioeconomic status (SES) populations. While interventions for the Global South exist, they often report limited success, par…
Planner and Executor: Collaboration between Discrete Diffusion And Autoregressive Models in Reasoning
Lina Berrayana, Ahmed Heakl, Muhammad Abdullah Sohail +3
Current autoregressive language models (ARMs) achieve high accuracy but require long token sequences, making them costly. Discrete diffusion language models (DDLMs) enable parallel…
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
Rocktim Jyoti Das, Harsh Singh, Diana Turmakhan +5
Scaling data and models has played a pivotal role in the remarkable progress of computer vision and language. Inspired by these domains, recent efforts in robotics have similarly f…