<  Back to the Polytechnique Montréal portal

LLMs and Stack Overflow discussions: reliability, impact, and challenges

Leuson Mario Pedro Da Silva, Jordan S. A. M. HI and Foutse Khomh

Dataset (2025)

Open Acess document at official publisher
An external link is available for this item
Show abstract
Hide abstract

Abstract

Since the release of ChatGPT in November 2022, the landscape of developer Q&A platforms, particularly Stack Overflow, has undergone significant changes. The ability of large language models (LLMs) to generate immediate, human-like responses to technical questions has started discussions on their potential to replace traditional Q&A platforms. This dataset was collected as part of an empirical study analyzing Stack Overflow questions and evaluating responses generated by ChatGPT and LLaMA.

The dataset supports research aimed at:

Assessing the reliability of LLM-generated answers and their potential long-term impact on platforms like Stack Overflow.

Tracking the evolution of user engagement with Stack Overflow post-ChatGPT’s release. Comparing the performance of ChatGPT and LLaMA across different topics.

Supplementary Material:
Department: Department of Computer Engineering and Software Engineering
PolyPublie URL: https://publications.polymtl.ca/64422/
Source: Zenodo
DOI: 10.5281/zenodo.15086541
Other DOIs related to this document: 10.5281/zenodo.15086542
Official URL: https://doi.org/10.5281/zenodo.15086541
Date Deposited: 04 Apr 2025 16:54
Last Modified: 30 Jan 2026 09:19
Cite in APA 7: Da Silva, L. M. P., S. A. M. HI, J., & Khomh, F. (2025). LLMs and Stack Overflow discussions: reliability, impact, and challenges [Dataset]. Zenodo. https://doi.org/10.5281/zenodo.15086541

Statistics

Dimensions

Repository Staff Only

View Item View Item