6 papers
Multi-Granular Node Pruning for Causal Circuit Discovery
Muhammad Umair Haider, Hammad Rizwan, Hassan Sajjad +1
Circuit discovery aims to identify minimal subnetworks that are responsible for specific behaviors in large language models (LLMs). Existing approaches primarily rely on iterative…
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
Moshood A. Fakorede, Krishna Upadhyay, A. B. Siddique +1
Large language models (LLMs) have shown strong performance on automated software engineering tasks, yet existing benchmarks focus primarily on library-style repositories, leaving m…
Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution
Muhammad Umair Haider, Hammad Rizwan, Hassan Sajjad +2
Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for model interpretation and control. We sy…
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
Daniel Liu, Krishna Upadhyay, Vinaik Chhetri +2
The rapid emergence of multi-agent AI systems (MAS), including LangChain, CrewAI, and AutoGen, has shaped how large language model (LLM) applications are developed and orchestrated…
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
Vinaik Chhetri, Krishna Upadhyay, A. B. Siddique +1
Artificial Intelligence (AI)-powered features have rapidly proliferated across mobile apps in various domains, including productivity, education, entertainment, and creativity. How…
Analyzing the Evolution and Maintenance of Quantum Software Repositories
Krishna Upadhyay, Vinaik Chhetri, A. B. Siddique +1
Quantum computing is rapidly advancing, but quantum software development faces significant challenges, including a steep learning curve, high hardware error rates, and a lack of ma…