5 papers · 1 filter
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
Moshood A. Fakorede, Krishna Upadhyay, A. B. Siddique +1
Large language models (LLMs) have shown strong performance on automated software engineering tasks, yet existing benchmarks focus primarily on library-style repositories, leaving m…
Understanding Robustness of Model Editing in Code LLMs
Vinaik Chhetri, Moghis Fereidouni, A. B Siddique +1
Large language models (LLMs) for code are increasingly used in software development, but they remain static after pretraining while APIs and software libraries continue to evolve.…
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
Daniel Liu, Krishna Upadhyay, Vinaik Chhetri +2
The rapid emergence of multi-agent AI systems (MAS), including LangChain, CrewAI, and AutoGen, has shaped how large language model (LLM) applications are developed and orchestrated…
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
Vinaik Chhetri, Krishna Upadhyay, A. B. Siddique +1
Artificial Intelligence (AI)-powered features have rapidly proliferated across mobile apps in various domains, including productivity, education, entertainment, and creativity. How…
Analyzing the Evolution and Maintenance of Quantum Software Repositories
Krishna Upadhyay, Vinaik Chhetri, A. B. Siddique +1
Quantum computing is rapidly advancing, but quantum software development faces significant challenges, including a steep learning curve, high hardware error rates, and a lack of ma…