<?xml version="1.0" encoding="UTF-8"?>
<TEI xmlns="http://www.tei-c.org/ns/1.0">
    <teiHeader>
        <fileDesc>
            <titleStmt>
                <title type="full">
                    <title type="main"><hi rend="italic">Open Stemmata</hi>: A Digital Collection of
                        Textual Genealogies</title>
                </title>
                <author>
                    <persName>
                        <surname>Camps</surname>
                        <forename>Jean-Baptiste</forename>
                    </persName>
                    <affiliation>École nationale des chartes | PSL, France</affiliation>
                    <email>jean-baptiste.camps@chartes.psl.eu</email>
                </author>
                <author>
                    <persName>
                        <surname>Gabay</surname>
                        <forename>Simon</forename>
                    </persName>
                    <affiliation>Université de Genève</affiliation>
                    <email>simon.gabay@unige.ch</email>
                </author>
                <author>
                    <persName>
                        <surname>Riva</surname>
                        <forename>Gustavo</forename>
                    </persName>
                    <affiliation>Universität Heidelberg</affiliation>
                    <email>gustavo.fernandez.riva@uni-heidelberg.de</email>
                </author>
            </titleStmt>
            <editionStmt>
                <edition>
                    <date>2021-06-15T11:06:00.869970060</date>
                </edition>
            </editionStmt>
            <publicationStmt>
                <publisher>Elisabeth Burr, University of Leipzig</publisher>
                <address>
                    <addrLine>Beethovenstr. 15</addrLine>
                    <addrLine>04107 Leipzig</addrLine>
                    <addrLine>Germany</addrLine>
                    <addrLine>Elisabeth Burr</addrLine>
                </address>
            </publicationStmt>
            <sourceDesc>
                <p>Converted from an OASIS Open Document</p>
            </sourceDesc>
        </fileDesc>
        <encodingDesc>
            <appInfo>
                <application ident="DHCONVALIDATOR" version="1.22">
                    <label>DHConvalidator</label>
                </application>
            </appInfo>
        </encodingDesc>
        <profileDesc>
            <textClass>
                <keywords scheme="ConfTool" n="category">
                    <term>Paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="subcategory">
                    <term>Short Paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="keywords">
                    <term>stemmatology</term>
                    <term>philology</term>
                </keywords>
                <keywords scheme="ConfTool" n="topics">
                    <term>Gathering</term>
                    <term>Programming</term>
                    <term>Annotating</term>
                    <term>Network Analysis</term>
                    <term>Visualization</term>
                    <term>Contextualizing</term>
                    <term>Modeling</term>
                    <term>Organizing</term>
                    <term>Collaboration</term>
                    <term>Crowdsourcing</term>
                    <term>Sharing</term>
                    <term>Images</term>
                    <term>Metadata</term>
                    <term>Manuscript</term>
                    <term>Data</term>
                    <term>not applicable</term>
                    <term>English</term>
                </keywords>
            </textClass>
        </profileDesc>
    </teiHeader>
    <text>
        <body>
            <div type="div1" rend="DH-Heading1">
                <head> Gathering stemmata
                </head>
                <p>
                    <hi rend="italic">Stemma codicum</hi> is the genealogical tree of the
                    manuscripts of a given text. More precisely, it is a tool of textual criticism
                    that represents the relationships between all the witnesses of a specific work
                    (Duval 2015; Roelli 2020) under the form of a tree, or, in the case of
                    contamination, a directed acyclic graph (Andrews / Macé 2012). The first
                    stemmata were drawn in the 1830’s, with at least one antecedent in the 18th
                    century, while the method to build them progressively took form during the 19th
                    century and is best called the “common errors” method (Camps / Cafiero 2014). </p>
                <p>In a stemma, the relationships between the witnesses and the hypothetical
                    necessary (lost) nodes are represented with a tree-like structure (cf. fig. 1).
                    The original purpose of the stemma is to allow for the reconstruction of the
                    archetype or the original text, as conceived by its author, but stemmata are
                    also used to study the transmission and the reception of works over centuries
                    (Marshall / Leighton Durham 1998), though this has sometimes spurred some debate
                    (Varvaro 2010; Croenen 2010). </p>
                <p>Historically, (personal) collections of stemmata have been used in epistemological debates on the common error methods. Joseph Bedier reports that he took the habit of tracing the stemmata he encountered, and built a collection of 110 stemmata, out of which he construed his remarks on the bifidity, from which he derived his criticism of the common error methods (Bédier 1928). To falsify or confirm Bédier’s claims, other scholars have replicated his endeavour, building their own collection (Shepard 1930; Castellani 1957; Haugen 2015) but, as far as we noticed, there is yet no publicly available digital collection of stemmata.</p>
                <p>A publicly available collection of stemmata would have a great interest, because
                    controversies on their shape are at the heart of the great philological debate
                    of the 20th century, at least for Romance Philology, where Bédier claims led
                    some philologists to renounce the common error method (e.g. in France), while it
                    gave other philological schools the necessary impulse to try and refine it
                    (e.g., in Italy, see Trovato 2014) <note xml:id="ftn1" place="foot" n="1">This
                        even led some philologists to suggest alternatives to the tree-like
                        structure, such as Venn diagram, unrooted graphs…(Carsten-Peust 2012). In
                        the case of some very complex traditions, some editors renounce to draw a
                        stemma (Perugi 2000); comp. to (Zufferey 2007).</note>. The development of
                    computational philology has been a good opportunity to think anew this old art
                    (Hoenen 2020). </p>
                <p>Moreover, study on the shapes of stemmata could prove interesting in many kinds of analysis, be it on the dissemination, reception and history of texts, or to be compared to more theoretical models.</p>
                <p>All of this triggers the need for a digital collections of stemmata, for mining and testing purposes. This paper presents an attempt to answer this need. We therefore propose to collect all available stemmata and encode them (cf. fig. 1a). Chronological and linguistic boundaries are very open: at least any European language is accepted, from every time period, even if we will focus first on Western Medieval languages (esp. French, Occitan, Italian, German, Spanish and of course Latin).</p>
             
                    <figure>
                        <graphic url="Pictures/349fc3ed952a34dca2704efb1ad1d0d6.jpg"/>
                    </figure>
                <p>Figure 1a: Segre’s original stemma</p>
              
                    <figure>
                        <graphic url="Pictures/8a3ad89a79ffbe786f28af80cf73cec5.png"/>
                    </figure>
                <p>Figure 2b: Graph visualisation of Segre’s stemma</p>
            </div>
            <div type="div1" rend="DH-Heading1">
                <head> Information, modelling and production
                </head>
                <p>
                    <hi rend="italic">Open Stemmata</hi> is a collaborative project, where researchers can participate by sharing stemmata that they encode. Guidelines have been published to help volunteers (OpenStemmata 2021), who are required to provide three documents for each stemma:
                </p>
                <list type="unordered">
                    <item>Metadata as a .txt file (such as source of the stemma and researcher responsible for the encoding).</item>
                    <item>A stemma encoded in the DOT format (Graphviz 2021), as a .gv file</item>
                    <item>A picture of the stemma, as a cropped PNG.</item>
                </list>
                <p>Metedata are produced with an online form
                    <note xml:id="ftn2" place="foot" n="2">Available on: https://openstemmata.github.io/document-your-stemma.html.</note> to ensure the homogeneity and the completeness of the data. Minimal standardisation is offered via VIAF (OCLC 2020) for author and work, and ORCID (ORCID 2021) for submission contributor.
                </p>
                <p>A stemma being a graph, we have decided to encode data with the DOT language
                    (Graphviz 2021) (fig. 3) which provides all the subtlety required and has a
                    simple syntax that enables easy manual encoding. Hypothetical nodes are
                    identified with “grey” as the value of the color attribute, contamination as
                    dashed lines. We suggest the use of an online graphviz editor (“Edotor,” n.d.),
                    but many other options are available.</p>
                <p>
                    <figure>
                        <graphic url="Pictures/0a7eae3367dd477a0fe64c861d1f0428.png"/>
                    </figure>
                </p>
                <p>Figure 3: Example of graph and its encoding in DOT language </p>
                <p>digraph {</p>
                <p> omega[label="Ω", color="grey"];</p>
                <p> omega -&gt;A;</p>
                <p> omega -&gt; 1;</p>
                <p> 1 -&gt; B;</p>
                <p> 1 -&gt; C;</p>
                <p> 1[label="", color="grey"];</p>
                <p> A -&gt; B [style="dashed", dir=none];</p>
                <p>}</p>
                <p>On top of the online form for the metadata, additional steps are in place to ensure data quality, documentation and re-usability.</p>
                <list type="unordered">
                    <item>On the one hand, data is shared, published and distributed via GitHub. It is compulsory to add stemmata with pull requests, which have to be validated by a member of the organisation. Data structure is then controlled through continuous integration and tests, but also by human inspection and quality control.</item>
                    <item>On the other hand, information is stored in a masterfile encoded according to the TEI Guidelines (TEI Consortium 2020): metadata in the teiHeader and the graph in the body. It allows basic interoperability and sustainability of our digital collection, but also a good documentation of the encoding choices (with a dedicated ODD file and schema).</item>
                    <item>Conversion scripts into GraphML, a common exchange format for graph and network analysis, as well as instructions on how to use them, will be available in the repository, in order to facilitate the re-use of the data for other researchers.</item>
                </list>
            </div>
            <div type="div1">
                <head> Present and Future of OpenStemmata
                </head>
                <p>At present, there are roughly 50 stemmata available in the database, most of them
                    concerning Old French traditions. In the future, we hope to have a coverage as
                    exhaustive as possible. Such data will allow many kinds of analyses on the shape
                    of stemmata, including comparative or specific analyses for different types of
                    traditions (fig. 4).</p>
                <p>
                    <figure>
                        <graphic url="Pictures/113ea205b1527b231f7b452df9dbd541.png"/>
                    </figure>
                </p>
                <p>Figure 4: Distribution of nodes (witnesses and hypothetical nodes) in a collection of Old French Epic stemmata (
                    <hi rend="italic">Chansons de geste</hi>)
                </p>
            </div>
        </body>
        <back>
            <div type="bibliogr">
                <listBibl>
                    <head>Bibliography</head>
                    <bibl>
                        <hi rend="bold">Andrews, Tara</hi> / <hi rend="bold">Macé, Caroline </hi>
                        (2012):. “Trees of Texts: Models and Methods for an Updated Theory of
                        Medieval Text Stemmatology Digital Humanities 2012”, in: <hi rend="italic"
                            >DH 2012</hi>. Hamburg &lt;<ref
                            target="http://www.dh2012.uni-hamburg.de/conference/programme/abstracts/trees-of-texts-models-and-methods-for-an-updated-theory-of-medieval-text-stemmatology.1.html"
                            >http://www.dh2012.uni-hamburg.de/conference/programme/abstracts/trees-of-texts-models-and-methods-for-an-updated-theory-of-medieval-text-stemmatology.1.html</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Bédier, Joseph</hi> (1928): “La Tradition Manuscrite Du Lai
                        de L’ombre : Réflexions Sur L’art d’éditer Les Anciens Textes”, in: <hi
                            rend="italic">Romania</hi> 54: 161–196, 321–356 DOI <ref
                            target="https://doi.org/10.3406/roma.1928.4345"
                            >https://doi.org/10.3406/roma.1928.4345</ref>. </bibl>
                    <bibl><hi rend="bold">Camps, Jean-Baptiste</hi> / <hi rend="bold">Cafiero,
                            Florian</hi>(2014): “Genealogical Variant Locations and Simplified
                        Stemma: A Test Case”, in: <hi rend="italic">Analysis of Ancient and Medieval
                            Texts and Manuscripts: Digital Approaches</hi> (edited by Tara Andrews
                        and Caroline Macé) (= Lectio 1). Turnhout: Brepols 69–93 DOI: <ref
                            target="https://doi.org/10.1484/M.LECTIO-EB.5.102565"
                            >https://doi.org/10.1484/M.LECTIO-EB.5.102565</ref>. </bibl>
                    <bibl>
                        <hi rend="bold">Carsten-Peust, Konstanz </hi>(2012): “The Stemma of the
                        Story of Sinuhe. Or: How to Use an Unrooted Phylogenetic Tree in Textual
                        Criticism”, in: <hi rend="italic">Lingua Aegyptia</hi> 20: 209–220 DOI: <ref
                            target="https://doi.org/10.11588/propylaeumdok.00002543"
                            >https://doi.org/10.11588/propylaeumdok.00002543</ref>. </bibl>
                    <bibl>
                        <hi rend="bold">Castellani, Arrigo </hi>(1957): <hi rend="italic">Bédier
                            Avait-Il Raison?: La Méthode de Lachmann Dans Les éditions de Textes Du
                            Moyen Age</hi> (= Discours Universitaires, Nouvelle Série / Freiburger
                        Universitätsreden, Neue Folge 20). Freiburg / Schweiz : Universitätsverlag /
                        Fribourg: Suisse. Éditions Universitaires. </bibl>
                    <bibl>
                        <hi rend="bold">Croenen, Godfried </hi>(2010): “Stemmata, Philology and
                        Textual History: A Response to Alberto Varvaro”, in: <hi rend="italic"
                            >Medioevo Romanzo</hi> 34: 422–426. </bibl>
                    <bibl>
                        <hi rend="bold">Duval, Frédéric</hi> (2015): <hi rend="italic">Les Mots de
                            L’édition de Textes</hi>. Magister. Paris: École nationale des chartes. </bibl>
                    <bibl><hi rend="bold">Edotor</hi> ( n.d.) &lt;<ref target="https://edotor.net/"
                            >https://edotor.net/</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Graphviz</hi> (2021): "The Dot Language:
                            Documentation" &lt;<ref
                                target="https://www.graphviz.org/doc/info/lang.html">https://www.graphviz.org/doc/info/lang.html</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Haugen, Odd Einar </hi>(2015): “The Silva Portentosa of
                        Stemmatology: Bifurcation in the Recension of Old Norse Manuscripts”, in:
                            <hi rend="italic">Digital Scholarship in the Humanities</hi> 31, 3:
                        594–610 DOI: <ref target="https://doi.org/10.1093/llc/fqv002">https://doi.org/10.1093/llc/fqv002</ref>. </bibl>
                    <bibl>
                        <hi rend="bold">Hoenen, Armin </hi>(2020): “History of Computer-Assisted
                        Stemmatology”, in: <hi rend="italic">Handbook of Stemmatology: History,
                            Methodology, Digital Approaches</hi>. Berlin: De Gruyter 294–303 DOI:
                        <ref target="https://doi.org/10.1515/9783110684384-006">https://doi.org/10.1515/9783110684384-006</ref>. </bibl>
                    <bibl>
                        <hi rend="bold">Marshall, Peter K. / Leighton Durham, Reynolds </hi>(1998):
                            <hi rend="italic">Texts and Transmission: A Survey of the Latin
                            Classics</hi>. Oxford: Clarendon Press. </bibl>
                    <bibl>
                        <hi rend="bold">OCLC</hi> (2020): “Virtual International Authority File
                        (VIAF)” &lt;<ref target="https://viaf.org/">https://viaf.org/</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">OpenStemmata </hi>(2021): "Guidelines" &lt;<ref
                            target="https://openstemmata.github.io/guidelines.html">https://openstemmata.github.io/guidelines.html</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">ORCID </hi>(2021): “Open Researcher and Contributor ID” &lt;<ref target="https://orcid.org/">https://orcid.org/</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Perugi, Maurizio</hi> (ed.) (2000): <hi rend="italic">La Vie
                            de Saint Alexis</hi> (= Textes Littéraires Français 529). Genève: Droz. </bibl>
                    <bibl>
                        <hi rend="bold">Roelli, Philipp </hi>(2020): <hi rend="italic">Handbook of
                            Stemmatology: History, Methodology, Digital Approaches</hi>. Berlin: De
                        Gruyter DOI: <ref target="https://doi.org/10.1515/9783110684384">https://doi.org/10.1515/9783110684384</ref>. </bibl>
                    <bibl>
                        <hi rend="bold">Shepard, William P. </hi>(1930): “Recent Theories of Textual
                        Criticism”, in: <hi rend="italic">Modern Philology</hi> 28, 2: 129–141. </bibl>
                    <bibl>
                        <hi rend="bold">TEI Consortium </hi>(2020): <hi rend="italic">TEI P5:
                            Guidelines for Electronic Text Encoding and Interchange</hi> (version
                        Version 4.1.0) &lt;<ref target="https://tei-c.org/Guidelines/">https://tei-c.org/Guidelines/</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Trovato, Paolo </hi>(2014): “Bédier’s Contribution to the
                        Accomplishment of Stemmatic Method: An Italian Perspective”, in: <hi
                            rend="italic">Textual Cultures</hi> 9, 1: 160–176. </bibl>
                    <bibl>
                        <hi rend="bold">Varvaro, Alberto </hi>(2010): “Un Nuovo Studio Sulla
                        Tradizione Delle Chroniques Di Jean Froissart”, in: <hi rend="italic"
                            >Medioevo Romanzo</hi> 34: 145–152. </bibl>
                    <bibl>
                        <hi rend="bold">Zufferey, François </hi>(2007): “La Tradition Manuscrite Du
                        Saint Alexis Primitif”, in: <hi rend="italic">Romania</hi> 125, 497: 1–45 DOI: <ref target="https://doi.org/10.3406/roma.2007.1387">https://doi.org/10.3406/roma.2007.1387</ref>. </bibl>
                </listBibl>
            </div>
        </back>
    </text>
</TEI>
