Skip to main navigation Skip to search Skip to main content

Referencing source code artifacts: A separate concern in software citation

  • Laboratoire de Probabilités et Modèles Aléatoires
  • Università Dell'Aquila

Research output: Contribution to journalArticlepeer-review

13 Citations (Scopus)

Abstract

Among the entities involved in software citation, software source code requires special attention due to the role it plays in ensuring scientific reproducibility. To reference source code, we need identifiers that are not only unique and persistent, but also support integrity checking intrinsically. Suitable identifiers must guarantee that denoted objects will always stay the same, without relying on external third parties and administrative processes. We analyze the role of identifiers for digital objects, whose properties are different from, and complementary to, those of the various digital identifiers of objects that are today popular building blocks of software and data citation toolchains. We argue that both kinds of identifiers are needed and detail the syntax, semantics, and practical implementation of the persistent identifiers adopted by the Software Heritage project to reference billions of software source code artifacts such as source code files, directories, and commits.

Original languageEnglish
Article number8946737
Pages (from-to)33-43
Number of pages11
JournalComputing in Science and Engineering
Volume22
Issue number2
DOIs
Publication statusPublished - 1 Mar 2020
Externally publishedYes

Fingerprint

Dive into the research topics of 'Referencing source code artifacts: A separate concern in software citation'. Together they form a unique fingerprint.

Cite this