Merchantry Knowledge

evidence / stable release

[[edit](/w/index.php?title=Provenance&action=edit&section=11 "Edit section: Computer science")] Within [computer science](//en.wikipedia.org/wiki/Computer_science "Computer science"), [informatics](//en.wikipedia.org/wiki/Information_science "Information science") uses the term "provenance"[[54]](#cite_note-54) to mean the [lineage of data](//en.wikipedia.org/wiki/Data_lineage "Data lineage"), as per data provenance, with research in the last decade extending the conceptual model of causality and relation to include processes that act on data and agents that are responsible for those processes. See, for example, the proceedings of the International Provenance Annotation Workshop (IPAW)[[55]](#cite_note-55) and Theory and Practice of Provenance (TaPP).[[56]](#cite_note-56)[Semantic web](//en.wikipedia.org/wiki/Semantic_web "Semantic web") standards bodies, including the [World Wide Web Consortium](//en.wikipedia.org/wiki/World_Wide_Web_Consortium "World Wide Web Consortium") in 2014, have ratified a standard data model for provenance representation known as PROV[[57]](#cite_note-57) which draws from many of the better-known provenance representation systems that preceded it, such as the [Proof Markup Language](//en.wikipedia.org/wiki/Proof_Markup_Language "Proof Markup Language") and the Open Provenance Model.[[58]](#cite_note-58) Interoperability is a design goal of most recent computer science provenance theories and models, for example the Open Provenance Model (OPM) 2008 generation workshop aimed at "establishing inter-operability of systems" through information exchange agreements.[[59]](#cite_note-59) Data models and serialisation formats for delivering provenance information typically reuse existing metadata models where possible to enable this. Both the OPM Vocabulary[[60]](#cite_note-60) and the PROV Ontology[[61]](#cite_note-61) make extensive use of metadata models such as [Dublin Core](//en.wikipedia.org/wiki/Dublin_Core "Dublin Core") and [Semantic Web](//en.wikipedia.org/wiki/Semantic_Web "Semantic Web") technologies such as the [Web Ontology Language](//en.wikipedia.org/wiki/Web_Ontology_Language "Web Ontology Language") (OWL). Current practice is to rely on the W3C PROV data model, OPM's successor.[[62]](#cite_note-62) There are several maintained and open-source provenance capture implementation at the operating system level such as CamFlow,[[63]](#cite_note-63)[[64]](#cite_note-64) Progger[[65]](#cite_note-:0-65) for Linux and MS Windows, and SPADE for Linux, [MS Windows](//en.wikipedia.org/wiki/Microsoft_Windows "Microsoft Windows"), and [MacOS](//en.wikipedia.org/wiki/MacOS "MacOS").[[66]](#cite_note-66) Operating system level provenance have gained interest in the security community notably to develop novel intrusion detection techniques.[[67]](#cite_note-67) Other implementations exist for specific programming and scripting languages, such as RDataTracker[[68]](#cite_note-68) for [R](//en.wikipedia.org/wiki/R_(programming_language) "R (programming language)"), and NoWorkflow[[69]](#cite_note-69) for [Python](//en.wikipedia.org/wiki/Python_(programming_language) "Python (programming language)").

unit:49b82edf77de8e07e792:13a5670cc77404aa14b8:1:450934ed6b20d48fc35f ยท release release:edition:knowledge-systems:7aaba55a11d29659

Canonical record

[[edit](/w/index.php?title=Provenance&action=edit&section=11 "Edit section: Computer science")] Within [computer science](//en.wikipedia.org/wiki/Computer_science "Computer science"), [informatics](//en.wikipedia.org/wiki/Information_science "Information science") uses the term "provenance"[[54]](#cite_note-54) to mean the [lineage of data](//en.wikipedia.org/wiki/Data_lineage "Data lineage"), as per data provenance, with research in the last decade extending the conceptual model of causality and relation to include processes that act on data and agents that are responsible for those processes. See, for example, the proceedings of the International Provenance Annotation Workshop (IPAW)[[55]](#cite_note-55) and Theory and Practice of Provenance (TaPP).[[56]](#cite_note-56)[Semantic web](//en.wikipedia.org/wiki/Semantic_web "Semantic web") standards bodies, including the [World Wide Web Consortium](//en.wikipedia.org/wiki/World_Wide_Web_Consortium "World Wide Web Consortium") in 2014, have ratified a standard data model for provenance representation known as PROV[[57]](#cite_note-57) which draws from many of the better-known provenance representation systems that preceded it, such as the [Proof Markup Language](//en.wikipedia.org/wiki/Proof_Markup_Language "Proof Markup Language") and the Open Provenance Model.[[58]](#cite_note-58) Interoperability is a design goal of most recent computer science provenance theories and models, for example the Open Provenance Model (OPM) 2008 generation workshop aimed at "establishing inter-operability of systems" through information exchange agreements.[[59]](#cite_note-59) Data models and serialisation formats for delivering provenance information typically reuse existing metadata models where possible to enable this. Both the OPM Vocabulary[[60]](#cite_note-60) and the PROV Ontology[[61]](#cite_note-61) make extensive use of metadata models such as [Dublin Core](//en.wikipedia.org/wiki/Dublin_Core "Dublin Core") and [Semantic Web](//en.wikipedia.org/wiki/Semantic_Web "Semantic Web") technologies such as the [Web Ontology Language](//en.wikipedia.org/wiki/Web_Ontology_Language "Web Ontology Language") (OWL). Current practice is to rely on the W3C PROV data model, OPM's successor.[[62]](#cite_note-62) There are several maintained and open-source provenance capture implementation at the operating system level such as CamFlow,[[63]](#cite_note-63)[[64]](#cite_note-64) Progger[[65]](#cite_note-:0-65) for Linux and MS Windows, and SPADE for Linux, [MS Windows](//en.wikipedia.org/wiki/Microsoft_Windows "Microsoft Windows"), and [MacOS](//en.wikipedia.org/wiki/MacOS "MacOS").[[66]](#cite_note-66) Operating system level provenance have gained interest in the security community notably to develop novel intrusion detection techniques.[[67]](#cite_note-67) Other implementations exist for specific programming and scripting languages, such as RDataTracker[[68]](#cite_note-68) for [R](//en.wikipedia.org/wiki/R_(programming_language) "R (programming language)"), and NoWorkflow[[69]](#cite_note-69) for [Python](//en.wikipedia.org/wiki/Python_(programming_language) "Python (programming language)").

Structured record
{
  "evidence_unit_id": "unit:49b82edf77de8e07e792:13a5670cc77404aa14b8:1:450934ed6b20d48fc35f",
  "artifact_id": "artifact:49b82edf77de8e07e792:6edc1142d1044e7186b5",
  "text": "[[edit](/w/index.php?title=Provenance&action=edit&section=11 \"Edit section: Computer science\")]\n\nWithin [computer science](//en.wikipedia.org/wiki/Computer_science \"Computer science\"), [informatics](//en.wikipedia.org/wiki/Information_science \"Information science\") uses the term \"provenance\"[[54]](#cite_note-54) to mean the [lineage of data](//en.wikipedia.org/wiki/Data_lineage \"Data lineage\"), as per data provenance, with research in the last decade extending the conceptual model of causality and relation to include processes that act on data and agents that are responsible for those processes. See, for example, the proceedings of the International Provenance Annotation Workshop (IPAW)[[55]](#cite_note-55) and Theory and Practice of Provenance (TaPP).[[56]](#cite_note-56)[Semantic web](//en.wikipedia.org/wiki/Semantic_web \"Semantic web\") standards bodies, including the [World Wide Web Consortium](//en.wikipedia.org/wiki/World_Wide_Web_Consortium \"World Wide Web Consortium\") in 2014, have ratified a standard data model for provenance representation known as PROV[[57]](#cite_note-57) which draws from many of the better-known provenance representation systems that preceded it, such as the [Proof Markup Language](//en.wikipedia.org/wiki/Proof_Markup_Language \"Proof Markup Language\") and the Open Provenance Model.[[58]](#cite_note-58)\n\nInteroperability is a design goal of most recent computer science provenance theories and models, for example the Open Provenance Model (OPM) 2008 generation workshop aimed at \"establishing inter-operability of systems\" through information exchange agreements.[[59]](#cite_note-59) Data models and serialisation formats for delivering provenance information typically reuse existing metadata models where possible to enable this. Both the OPM Vocabulary[[60]](#cite_note-60) and the PROV Ontology[[61]](#cite_note-61) make extensive use of metadata models such as [Dublin Core](//en.wikipedia.org/wiki/Dublin_Core \"Dublin Core\") and [Semantic Web](//en.wikipedia.org/wiki/Semantic_Web \"Semantic Web\") technologies such as the [Web Ontology Language](//en.wikipedia.org/wiki/Web_Ontology_Language \"Web Ontology Language\") (OWL). Current practice is to rely on the W3C PROV data model, OPM's successor.[[62]](#cite_note-62)\n\nThere are several maintained and open-source provenance capture implementation at the operating system level such as CamFlow,[[63]](#cite_note-63)[[64]](#cite_note-64) Progger[[65]](#cite_note-:0-65) for Linux and MS Windows, and SPADE for Linux, [MS Windows](//en.wikipedia.org/wiki/Microsoft_Windows \"Microsoft Windows\"), and [MacOS](//en.wikipedia.org/wiki/MacOS \"MacOS\").[[66]](#cite_note-66) Operating system level provenance have gained interest in the security community notably to develop novel intrusion detection techniques.[[67]](#cite_note-67) Other implementations exist for specific programming and scripting languages, such as RDataTracker[[68]](#cite_note-68) for [R](//en.wikipedia.org/wiki/R_(programming_language) \"R (programming language)\"), and NoWorkflow[[69]](#cite_note-69) for [Python](//en.wikipedia.org/wiki/Python_(programming_language) \"Python (programming language)\").",
  "access_class": "public",
  "heading": "Computer science",
  "observed_at": "2026-07-18T21:40:59.214Z",
  "state": "active"
}