1 citations · 1 across the 7 of their papers we have counts for
28 papers
From Business Requirements to Test Assertions: Evaluating LLM-Generated Oracles on Real Bugs
Tiancheng Ma, Nasir U. Eisty
The oracle problem (determining the correct expected outcome for a test) remains a major bottleneck in automated testing, and is increasingly relevant as non-experts rely on AI-gen…
Self-Admitted Technical Debt in Scientific Software: Prioritization, Sentiment, and Propagation Across Artifacts
Eric L. Melin, Nasir U. Eisty, Gregory R. Watson +1
Self-admitted technical debt (SATD) impairs scientific software (SSW), yet its prioritization, sentiment, persistence, and propagation remains underexplored. Understanding how SSW…
Thinking Out Loud: Real-Time Deception Monitoring in Asymmetric LLM Negotiations
Nolan Coffey, Faithful Odoi, Makenzie Johnson +1
As LLM-based agents are increasingly deployed to negotiate, delegate, or transact on a user's behalf, software pipelines need runtime mechanisms to verify that an agent's stated in…
SentTrack: Sentiment-Driven Bottleneck Detection in GitHub Issue Repositories
Xinyu Hu, Ali Behbahani, Daniel Moon +2
Software engineering teams increasingly depend on GitHub issue threads to coordinate work, report bugs, and negotiate technical decisions, yet most repository health tools focus on…
LLM vs. Human Unit Tests: Fault Detection on Real Python Bugs
Phouvadeth Vathana, Prapti Bhatt, Rishi Patel +1
Large language models (LLMs) have shown considerable promise for automated unit test generation, yet their practical effectiveness relative to human-written tests remains poorly un…
Exploring Sustainability in Scientific Software through Code Quality & Test Coverage Metrics
Sheikh Md. Mushfiqur Rahman, Gregory R. Watson, Nasir U. Eisty
Context: Scientific open-source software (SciOSS) plays a foundational role in research and engineering, yet its long-term sustainability has often been overlooked and remains a si…