Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
110 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

GitHub Repositories with Links to Academic Papers: Public Access, Traceability, and Evolution (2004.00199v3)

Published 1 Apr 2020 in cs.SE and cs.DL

Abstract: Traceability between published scientific breakthroughs and their implementation is essential, especially in the case of open-source scientific software which implements bleeding-edge science in its code. However, aligning the link between GitHub repositories and academic papers can prove difficult, and the current practice of establishing and maintaining such links remains unknown. This paper investigates the role of academic paper references contained in these repositories. We conduct a large-scale study of 20 thousand GitHub repositories that make references to academic papers. We use a mixed-methods approach to identify public access, traceability and evolutionary aspects of the links. Although referencing a paper is not typical, we find that a vast majority of referenced academic papers are public access. These repositories tend to be affiliated with academic communities. More than half of the papers do not link back to any repository. We find that academic papers from top-tier SE venues are not likely to reference a repository, but when they do, they usually link to a GitHub software repository. In a network of arXiv papers and referenced repositories, we find that the most referenced papers are (i) highly-cited in academia and (ii) are referenced by repositories written in different programming languages.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (7)
  1. Supatsara Wattanakriengkrai (14 papers)
  2. Bodin Chinthanet (11 papers)
  3. Hideaki Hata (48 papers)
  4. Raula Gaikovina Kula (83 papers)
  5. Christoph Treude (137 papers)
  6. Jin Guo (42 papers)
  7. Kenichi Matsumoto (73 papers)
Citations (15)

Summary

We haven't generated a summary for this paper yet.