REVIEW 4 cited by
Transforming Software Development: Evaluating the Efficiency and Challenges of GitHub Copilot in Real-World Projects
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Generative AI technologies promise to transform the product development lifecycle. This study evaluates the efficiency gains, areas for improvement, and emerging challenges of using GitHub Copilot, an AI-powered coding assistant. We identified 15 software development tasks and assessed Copilot's benefits through real-world projects on large proprietary code bases. Our findings indicate significant reductions in developer toil, with up to 50% time saved in code documentation and autocompletion, and 30-40% in repetitive coding tasks, unit test generation, debugging, and pair programming. However, Copilot struggles with complex tasks, large functions, multiple files, and proprietary contexts, particularly with C/C++ code. We project a 33-36% time reduction for coding-related tasks in a cloud-first software development lifecycle. This study aims to quantify productivity improvements, identify underperforming scenarios, examine practical benefits and challenges, investigate performance variations across programming languages, and discuss emerging issues related to code quality, security, and developer experience.
Forward citations
Cited by 4 Pith papers
-
Do These Violent Delights Have Violent Ends? Measuring the Post-Merge Fate of Agentic Code
Post-merge, agentic code needs ~46–51% more corrective/bug-fix maintenance and introduces more security and dependency findings than human code, with higher burden in low-review projects.
-
Persistent Sparse Autoencoders: Learning Feature Timescales in Language Models
Persistent SAEs learn per-feature persistence coefficients from reconstruction, splitting features into fast local detectors and slow topic-tracking states that retain prompt-injection signals over long contexts.
-
Three-Phase Evaluation of AI-Assisted Software Development Life Cycle
In a sequential three-phase pilot, higher AI autonomy and AWS Kiro versus GitHub Copilot were associated with fewer development hours, higher RITM scores, and lower mental demand, with modestly higher frustration.
-
The Ground Is Shifting: A Reflection on the Foundations of Software Measurement
AI-generated commits and squash-merging violate the human-origin assumption underlying software measurement, and the field should run a large-scale AI-assisted replication agenda.
Discussion (0). Continue with ORCID to comment.