Skip to content
Preprint

Do Stack Overflow Answer Edits Occur Beyond Java? A Replication on Python and JavaScript

Aug 2026 · 1 citation · 17 references
Computer Science

TL;DR

The central findings of the original study generalise beyond Java, with the supply of candidate improvements considerably larger in both replication languages than in the original.

Abstract

Stack Overflow answers are continually revised by the community, and the edits made to their code snippets are a potential source of improvements for code that has been reused in open-source projects. A recent empirical study established this for Java, reporting that 16.11% of accepted Java answers are edited and that the resulting recommendations concentrate in highly popular GitHub projects. Whether that behaviour is a property of Stack Overflow or a property of the Java community has remained an open question. We replicate the study on Python and JavaScript, the two most widely used languages alongside Java, applying the same SOTorrent-based extraction pipeline, the same clone search tool, Siamese+, and the same project popularity criteria. Analysing 840,132 accepted Python answers and 1,144,185 accepted JavaScript answers, we find that 41.25% and 39.10% respectively have been edited at least once, roughly two and a half times the Java rate, while the number of revisions per edited answer is almost invariant across the three languages at 2.78, 2.68 and 2.82. Searching 100 GitHub projects per language, we find that the number of matched answer edits increases monotonically from low- to medium- to high-popularity projects in both languages, from 80 to 156 to 977 for Python and from 32 to 71 to 353 for JavaScript. The difference is statistically significant for Python but not for JavaScript. The central findings of the original study therefore generalise beyond Java, with the supply of candidate improvements considerably larger in both replication languages than in the original.

View source

Similar papers

Preprint Aug 2026

A Comprehensive Study of Native Code Bugs in Python Applications

The impact of Python applications has been evidenced by their widespread presence in some of the most impactful software domains, such as machine learning frameworks and scientific computing platforms. These applications often integrate native code components written in a lower-level programming language like C. This m...

Haoran Yang, Hai-Peng Cai · 0 citations
Book Open access Oct 2026

In-the-Wild JavaScript Dead-Code Elimination via HTTP Range Requests

JavaScript (JS) bloat is a major source of slow, data-inefficient web performance, as many sites ship large monolithic bundles with substantial unused code. Existing JS dead-code elimination techniques help but require server-side support and are therefore hard to deploy. This paper evaluates HTTP range requests as a c...

Ayush Pandey, Vladimir Sharkovski, Matteo Varvello et al. · 0 citations
Preprint Aug 2026

Revisiting Feedback-Driven LLM Code Repair: A Replication and Exploratory Java Extension

The results show that previous conclusions from Python may be sensitive to benchmark construction, feedback representation, and tooling ecosystem, motivating more controlled multilingual benchmarks, and confirm key trends under a partially controlled replication of LLMbased repair systems.

L. Lalonde, Wassim Keddache, Thomas Perron Touchette et al. · 0 citations
Review Open access Aug 2026

A Systematic Literature Review of Code Smell Detection Tools for JavaScript Systems

A systematic literature review of JavaScript code smell detection tools found that most tools use rule-based linting, which is efficient but struggles with complex architectural smells, and AI-driven detection is completely missing.

Saymon Souza, Felipe Ribeiro, Eduardo Fernandes et al. · 0 citations
Book Open access Sep 2026

Quantifying the Code-Size Overhead of eBPF JIT Compilation

eBPF allows user-defined programs to safely extend Linux kernel functionality at runtime, but its final machine code comes from a compilation pipeline that differs from native targets, and how efficient that pipeline is has no clear reference point. Our work constructs one: using the standard LLVM x86 backend as an app...

Hoang Duong, Hao Sun, Zhen-Dong Su · 0 citations
#natural language process... Preprint Sep 2026

Large Language Models for Programming: Actually Fixing or Reimplementing Incorrect Code?

Recent studies have shown that Large Language Models can effectively solve problems and fix bugs in diverse programming environments, including competitive programming. Existing approaches primarily evaluate LLM performance in problem solving or bug fixing independently, but do not explore the relationship between thes...

Alexandru Stefan Stoica, Traian Rebedea, M. Mihăescu · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.