Conference paper Open Access

PyDriller: Python Framework for Mining Software Repositories

Davide Spadini; Mauricio Aniche; Alberto Bacchelli

MARC21 XML Export

<?xml version='1.0' encoding='UTF-8'?>
<record xmlns="">
  <controlfield tag="005">20200120172040.0</controlfield>
  <controlfield tag="001">1327411</controlfield>
  <datafield tag="711" ind1=" " ind2=" ">
    <subfield code="g">ESEC/FSE</subfield>
    <subfield code="a">26th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering</subfield>
  <datafield tag="700" ind1=" " ind2=" ">
    <subfield code="u">Delft University of Technology</subfield>
    <subfield code="a">Mauricio Aniche</subfield>
  <datafield tag="700" ind1=" " ind2=" ">
    <subfield code="u">University of Zurich</subfield>
    <subfield code="0">(orcid)0000-0003-0193-6823</subfield>
    <subfield code="a">Alberto Bacchelli</subfield>
  <datafield tag="856" ind1="4" ind2=" ">
    <subfield code="s">443338</subfield>
    <subfield code="z">md5:70fee7b40f0163ea17dff5d4291615c8</subfield>
    <subfield code="u"></subfield>
  <datafield tag="542" ind1=" " ind2=" ">
    <subfield code="l">open</subfield>
  <datafield tag="260" ind1=" " ind2=" ">
    <subfield code="c">2018-08-03</subfield>
  <datafield tag="909" ind1="C" ind2="O">
    <subfield code="p">openaire</subfield>
    <subfield code="p">user-msr</subfield>
    <subfield code="o"></subfield>
  <datafield tag="100" ind1=" " ind2=" ">
    <subfield code="u">Delft University of Technology</subfield>
    <subfield code="0">(orcid)0000-0003-2997-1890</subfield>
    <subfield code="a">Davide Spadini</subfield>
  <datafield tag="245" ind1=" " ind2=" ">
    <subfield code="a">PyDriller: Python Framework for Mining Software Repositories</subfield>
  <datafield tag="980" ind1=" " ind2=" ">
    <subfield code="a">user-msr</subfield>
  <datafield tag="536" ind1=" " ind2=" ">
    <subfield code="c">642954</subfield>
    <subfield code="a">Software ENgineering in Enterprise Cloud Applications  systems</subfield>
  <datafield tag="536" ind1=" " ind2=" ">
    <subfield code="c">PP00P2_170529</subfield>
    <subfield code="a">Data-driven Contemporary Code Review</subfield>
  <datafield tag="540" ind1=" " ind2=" ">
    <subfield code="a">Other (Open)</subfield>
  <datafield tag="650" ind1="1" ind2="7">
    <subfield code="a">cc-by</subfield>
    <subfield code="2"></subfield>
  <datafield tag="520" ind1=" " ind2=" ">
    <subfield code="a">&lt;p&gt;Software repositories contain historical and valuable information about the overall development of software systems. Mining software repositories (MSR) is nowadays considered one of the most interesting growing fields within software engineering. MSR focuses on extracting and analyzing data available in software repositories to uncover interesting, useful, and actionable information about the system. Even though MSR plays an important role in software engineering research, few tools have been created and made public to support developers in extracting information from Git repository. In this paper, we present PyDriller, a Python Framework that eases the process of mining Git. We compare our tool against the state-of-the-art Python Framework GitPython, demonstrating that PyDriller can achieve the same results with, on average, 50% less LOC and significantly lower complexity.&lt;br&gt;
  <datafield tag="773" ind1=" " ind2=" ">
    <subfield code="n">doi</subfield>
    <subfield code="i">isVersionOf</subfield>
    <subfield code="a">10.5281/zenodo.1327410</subfield>
  <datafield tag="024" ind1=" " ind2=" ">
    <subfield code="a">10.5281/zenodo.1327411</subfield>
    <subfield code="2">doi</subfield>
  <datafield tag="980" ind1=" " ind2=" ">
    <subfield code="a">publication</subfield>
    <subfield code="b">conferencepaper</subfield>
All versions This version
Views 378378
Downloads 297297
Data volume 131.7 MB131.7 MB
Unique views 352352
Unique downloads 271271


Cite as