<?xml version="1.0" encoding="UTF-8"?>
<TEI xmlns="http://www.tei-c.org/ns/1.0">
    <teiHeader>
        <fileDesc>
            <titleStmt>
                <title>Periodicals in Motion: Textual Reuse and the Hebrew Journalistic Networks in the second half of the 19th century</title>
                <author>
                    <persName>
                        <surname>Segal</surname>
                        <forename>Zef</forename>
                    </persName>
                    <affiliation>The Open University of Israel, Israel</affiliation>
                    <email>zefsegal@gmail.com</email>
                </author>
            </titleStmt>
            <editionStmt>
                <edition>
                    <date>2021-09-17T14:42:00Z</date>
                </edition>
            </editionStmt>
            <publicationStmt>
                <publisher>Elisabeth Burr, University of Leipzig</publisher>
                <address>
                    <addrLine>Beethovenstr. 15</addrLine>
                    <addrLine>04107 Leipzig</addrLine>
                    <addrLine>Germany</addrLine>
                    <addrLine>Elisabeth Burr</addrLine>
                </address>
            </publicationStmt>
            <sourceDesc>
                <p>Converted from a Word document</p>
            </sourceDesc>
        </fileDesc>
        <encodingDesc>
            <appInfo>
                <application ident="DHCONVALIDATOR" version="1.22">
                    <label>DHConvalidator</label>
                </application>
            </appInfo>
        </encodingDesc>
        <profileDesc>
            <textClass>
                <keywords scheme="ConfTool" n="category">
                    <term>Paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="subcategory">
                    <term>Long paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="keywords">
                    <term>periodical studies</term>
                    <term>textual reuse</term>
                    <term>stylometry</term>
                    <term>Hebrew OCR</term>
                    <term>historical network analysis</term>
                </keywords>
                <keywords scheme="ConfTool" n="topics">
                    <term>DataRecognition</term>
                    <term>Discovering</term>
                    <term>Gathering</term>
                    <term>Writing</term>
                    <term>Network Analysis</term>
                    <term>Relational Analysis</term>
                    <term>Stylistic Analysis</term>
                    <term>Contextualizing</term>
                    <term>Theorizing</term>
                    <term>Text</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>English</term>
                </keywords>
            </textClass>
        </profileDesc>
    </teiHeader>
    <text>
        <body>
            <p>The second half of the nineteenth century saw the establishment of numerous Hebrew periodicals that would play a major role in constituting a modern Hebrew “Republic of letters” (see Figure 1). They provided a platform for a lively discourse reflecting diverse ideological, political, and cultural approaches. In the multilingual context of Jewish communities, Hebrew was not an obvious choice for a journal, but it had the advantage of bridging the geographical and cultural distances between individuals and communities spread all over the Jewish Diaspora (Bartal 1994; Blondheim 1997; Soffer 2004a; Bartal 2007; Soffer 2009; Beer-Marx 2017). For Jewish communities—far from their homeland, lacking a central political and economic leadership, and spread throughout the world—the Hebrew press functioned as a public-sphere (Penslar 2000). Issues of these journals found their way to Jewish communities all around the world, thousands of miles far from their place of publication.
            </p>
            <figure>
                <graphic n="1001" width="8.995833333333334cm" height="7.006166666666667cm" url="Pictures/bdeeaf9e7880b9cb71d1ca5573f421be.jpeg" rend="inline"/>
            </figure>
            <p>Figure 1: A three-dimensional map of the editorial locations of late nineteenth
                century Hebrew periodicals and their movement through time and space. The vertical
                axis represents the time of publication, and each vertical box represents a single
                periodical. Blue arrows reflect movement of the editorial locations of a certain
                periodical. Relocated periodicals are represented with a similar colored box in
                their new locations. The map was created in QGIS.</p>
            <p>Hebrew journals were varied in nature and degree of institutionalization but most of them were limited in human and financial resources. As such, the publication of letters from private writers in near and far Jewish communities, made up a major part of the early Hebrew weeklies. The special characteristics of the Hebrew journals—their perception of themselves as communal enterprises; the geographical and cultural distance between communities that imagined themselves part of the same entity; and the relatively small group of modern Hebrew writers at the time—take to the extreme the "network authorship" journalistic model (Cordell 2015). This model assumes textual circulation and composition that is communal rather than individual. This communal authorship finds its expression, among other things, in the reuse and re-printing of texts. This reuse of texts occurs for different reasons: it could be a legitimate acknowledged citation of a Hebrew news item or a citation of the same translated news article from foreign newspapers, but it could also be a result of intentional plagiarism (Segal 2019). The anonymity of many authors contributed to this re-use of text.</p>
            <p>While previous studies of Jewish journalistic networks used qualitative research methods (Bartal 1994; Soffer 2007; Kouts 2013, Beer-Marx 2017) such as discourse analysis, this study uses computational tools to provide a wider perspective on the phenomenon of Hebrew periodical networks. The corpus consists of five major Hebrew journals (HaTzfira, HaMagid, HaMelitz, HaLebanon, and Havazelet) published between 1874 and 1883. Recent digital approaches to the study of historical periodicals have shown such tools can provide a “distant reading”, which offers a new generalized approach to otherwise untraceable periodical networks (Murphy 2014).</p>
            <div type="div1" rend="DH-Heading1">
                <head>Digital approaches to journalism networks</head>
                <p>The growing academic field of periodical studies is a direct result of advances
                    in digital technology in the last two decades (Latham / Scholes 2006; DiCenzo
                    2015). Keyword-searchable digital archives and algorithmic mining tools have
                    become accessible to an increasing number of researchers, and have transformed
                    our view of journals from mere containers of discrete bits of information to
                    autonomous objects of analysis. The most important of these tools is network
                    analysis. The periodical’s complex and composite form "embodies the concept of
                    the network on both a material level (in the juxtapositions and interconnections
                    it generates between different texts) and on an institutional level (in the
                    collaboration between authors, editors, illustrators, publishers, and readers,
                    which goes into producing it)" (Fagg et al. 2013). Accordingly, newspapers and
                    journals offer researchers a perspective to engage with the question of how
                    social and intellectual connections are forged, furthered, and diffused within a
                    public sphere. The interest in periodical networks is seen in special issues
                    dedicated to the topic in <hi rend="italic">Victorian Periodicals Review</hi>
                    (2011), <hi rend="italic">American Periodicals</hi> (2013) and <hi rend="italic"
                        >The Journal of Modern Periodical Studies</hi> (2014). </p>
                <p>The network metaphor is primarily used as a model of two aspects of periodical culture. First, a single periodical forms an intertextual network of individual texts, authors, and titles (Murphy 2014; Segal forthcoming). Second, periodicals are rarely isolated and exist within a larger print culture. Contributors, editors, publishers, and readers form an important, though sometimes barely visible, social and institutional network connecting different periodicals (Colavizza et al. 2014; Cordell 2015). At the same time, this line drawn between a journal’s internal and external networks is extremely permeable and porous, since the contacts between internal and external "nodes" established in previous editions always form the basis for ongoing editorial decisions. Consequently, external and internal factors are only artificially separable over the course of a journal project.</p>
                <p>The main problem is in detecting the existence and structure of the networks. In
                    some cases, we can find a network of authors by recognizing recurring names
                    (Haberman 2008; Ehrlicher / Herzgsell 2016). However, during the nineteenth
                    century many articles lacked clear authorship, due to, among other things, a
                    culture of reprinting (Cordell 2015) or the use of pseudonyms (Unsworth / Morton
                    1981; Mintz 1995). Although this excludes effectively using name-recognition
                    methodology to uncover periodical networks, hints remain within the text. "In
                    some cases," state Smith et al. (2013), "we cannot directly observe network
                    links, or even get a census of network nodes, and yet we can still observe text
                    that provides evidence for social interactions." Gennete’s (1997) notion of
                    transtextuality – "all that sets the text in relationship, whether obvious or
                    concealed, with other texts" – is useful in understanding Smith's remark.
                    Instead of approaching networks through people and working outwards, we can
                    identify and analyze networks through published texts and their similarities in
                    exact phrasing and styles.</p>
            </div>
            <div type="div1" rend="DH-Heading1">
                <head>Methodologies</head>
                <div type="div2" rend="DH-Heading2">
                    <head>Textual reuse</head>
                    <p>One direct method of tracing transtextuality is to identify copied and quoted
                        texts within other texts. However, recycled text is often only a small
                        portion of a complete document and may also be significantly rephrased. As a
                        result, finding these reused fragments "has always been subject to the
                        limitations of human reading and recollection" (Olsen et. al 2011). However,
                        as Olsen et al. continue, "one tantalizing promise of emerging digital
                        libraries is that computer technology may augment the scholarly functions of
                        reading and recollection by identifying related passages in very large
                        collections." This vision is supported by developments in the field of
                        plagiarism detection (Brin et al. 1995; Lovepreet / Kumar 2020), as well as
                        tailor-made computer codes for specific digital humanities projects (Lee
                        2007; Smith et al. 2013; Ganascia et al. 2014). </p>
                    <p>Accordingly, this study utilizes a commercial software, <hi rend="italic"
                            >Originality</hi>, used by Israeli universities to check the originality
                        of academic work. The software analyzes documents based on a previous
                        corpus, identifying copied passages and sentences, and measuring the level
                        of originality. It provides a detailed report for every document, which
                        marks the reused text and its source. The results are astonishing and prove
                        the existence of extensive and varied forms of textual reuse. The corpus
                        consists of 27,348 journalistic items, including advertisements. Some form
                        of textual overlap was identified in 12,102 of these items (see Figure 2 for
                        network visualization), and the average time gap between an original text
                        and its reuse was a little less than seven months. </p>
                    <figure>
                        <graphic n="1002" width="10.033cm" height="10.2235cm" url="Pictures/1e3f99da2212033db25458aba25c58a9.png" rend="inline"/>
                    </figure>
                    <p>Figure 2: Textual reuse, visualized as a network. In this figure, each node represents a single journalistic item, and each link represents an overlap of sentences between a pair of items. The graph was created by Gephi.</p>
                </div>
                <div type="div2" rend="DH-Heading2">
                    <head>Stylometry</head>
                    <p>Text reuse indicates news flows and direct intellectual links. But a much
                        more nuanced connection between texts can be revealed by similarity in
                        style. Stylometry, the statistical analysis of a literary style, is used to
                        prove the authenticity of documents or to settle questions of authorial
                        identity in anonymous or disputed texts (Mosteller / Wallace 1963; <hi
                            rend="italic">Chance</hi> 2003). Much like our ability to recognize
                        individual voices, stylometrists attempt to evaluate the lexical richness of
                        texts and establish similarities and differences between authors. </p>
                    <p>This study uses stylo, a flexible R package for high-level analysis of writing style (Eder et al. 2016). Instead of using stylometric analysis for authorship detection, it is used as a classification method, differentiating between literary circles and journalistic genres (Stamatatos et. al 2000). This approach is based on the distribution of the most frequent words (MFW) within each textual item (articles or advertisements). In order to reduce the noise of the results, all non-Hebrew letters were omitted from the texts and all the textual files smaller than 500 bytes were omitted from the corpus. This approach was tested on a limited corpus of 6,175 journalistic items published in the five journals between 1882 and 1883. Our results show stylistic similarity between 49,236 pairs of items, 1,654 of which are meaningful similarities between journals. These similarities range from identical items or authors to similar tones and subjects. </p>
                    <p>The evaluation of the results was conducted in a mixed-methods approach: edge weights and delta-similarity were used to identify proximity between journalistic items, network analysis to identify clusters, and close reading of a smaller sample to develop a typology of similarities. </p>
                </div>
            </div>
            <div type="div1" rend="DH-Heading1">
                <head>Results</head>
                <p>The results prove the existence of a journalistic network of writers, information, and readers, among journals published in distant locations and different ideological orientations. They reflect an increasing number of inter-journal writers, announcements, advertisements, as well as on-going discussions. As it seems, publishers, writers, and advertisers were aware of the existence of these transnational connections. The rapid development of a transnational Jewish public sphere in the second half of the 19
                    <hi rend="superscript">th</hi> century is revealed in this analysis. These results shed light on the multiple affordances of applying textual reuse and stylometric analysis to a large corpus of Hebrew journals. Those relate to the significance of Hebrew journals in the production and maintenance of a Jewish national imagination; to the characteristics of the emerging Hebrew print culture; and to the general understanding of journalistic networks.
                </p>
                <p>This paper presents the research, its integration of multiple computational tools, and suggests a categorization of textual similarities and reuse: from linguistic conventions of the Hebrew language, such as reuse of biblical and rabbinic phrases, to outright plagiarism of articles.</p>
            </div>
        </body>
        <back>
            <div type="bibliogr">
                <listBibl>
                    <head>Bibliography</head>
                    <bibl>
                        <hi rend="bold">AA.VV. </hi>(2002): <hi rend="italic">Chance</hi> 16, 2. </bibl>
                    <bibl>
                        <hi rend="bold">AA.VV. </hi>(2011): <hi rend="italic">Victorian Networks and
                            the Periodical Press</hi>
                        <hi rend="bold">.</hi> Special Issue. <hi rend="italic">Victorian
                            Periodicals Review</hi> 44, 2. </bibl>
                    <bibl>
                        <hi rend="bold">AA.VV.</hi> (2013): <hi rend="italic">Networks and the
                            Nineteenth-Century Periodical</hi>. Special Issue. <hi rend="italic"
                            >American Periodicals</hi> 23, 2. </bibl>
                    <bibl>
                        <hi rend="bold">AA.VV</hi>. (2014): <hi rend="italic">Visualizing Periodical
                            Networks</hi>. Special Issue. <hi rend="italic"> The Journal of Modern
                            Periodical Studies</hi> 5. </bibl>
                    <bibl>
                        <hi rend="bold">Bartal, Israel</hi> (1994): “Herald and Informer to a Jewish
                        Man: Jewish Journalism as a Channel to Innovation”, in: <hi rend="italic"
                            >Katedra</hi> (in Hebrew) 71: 154-164. </bibl>
                    <bibl>
                        <hi rend="bold">Bartal, Israel</hi> (2007): “From ‘Kahal’ to Readers
                        Community”, in: Soffer, Oren (ed.): <hi rend="italic">There Is No Place for
                            Pilpul! HaTzfira Journal and the Modernization of the Socio-political
                            Discourse</hi>. Jerusalem: Mossad Bialik Press with the Center for
                        Research on the History and Culture of Polish Jewry at the Hebrew
                        University. [Hebrew]. </bibl>
                    <bibl>
                        <hi rend="bold">Beer-Marx, Roni</hi> (2017): <hi rend="italic">Fortresses of
                            Paper: The Newspaper HaLevanon and Jewish Orthodoxy</hi>. Jerusalem: The
                        Zalman Shazar Center for Jewish History. [Hebrew]. </bibl>
                    <bibl>
                        <hi rend="bold">Colavizza, Giovanni</hi> / <hi rend="bold">Infelise,
                                Mario<hi rend="italic"> / </hi>Kaplan, Frederic</hi> (2015):
                        “Mapping the Early Modern News Flow: An Enquiry by Robust Text Reuse
                        Detection”, in: Aiello, Luca M. / Mcfarland, Daniel (eds.): <hi
                            rend="italic"> Social Informatics 6th International Conference, Socinfo
                            2014, Barcelona, Spain November 11-13, 2014 Proceedings</hi>. Cham:
                        Springer 244-253. </bibl>
                    <bibl>
                        <hi rend="bold">Cordell, Ryan</hi> (2015): “Reprinting, Circulation, and the
                        Network Author in Antebellum Newspapers”, in: <hi rend="italic">American
                            Literary History</hi> 27, 3: 417-445. </bibl>
                    <bibl>
                        <hi rend="bold">Eder, Maciej</hi> / <hi rend="bold">Rybicki, Jan</hi> / <hi
                            rend="bold">Kestemont, Mike</hi> (2016): “Stylometry with R: A Package
                        for Computational Text Analysis”, in: <hi rend="italic">R journal</hi> 8, 1:
                        107-121. </bibl>
                    <bibl>
                        <hi rend="bold">Ehrlicher, Hanno</hi> / <hi rend="bold">Herzgsell,
                            Teresa</hi> (2016): “Magazines as Networks and their Digital
                        Visualization: Basic Methodological Considerations and First Application
                        Examples”, in: <hi rend="italic">Revistas Culturales</hi> &lt;<ref
                            target="https://www.revistas-culturales.de/de/buchseite/hanno-ehrlicher-teresa-herzgsell-zeitschriften-als-netzwerke-und-ihre-digitale"
                            >https://www.revistas-culturales.de/de/buchseite/hanno-ehrlicher-teresa-herzgsell-zeitschriften-als-netzwerke-und-ihre-digitale</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Fagg, John</hi> / <hi rend="bold">Pethers, Matthew</hi> /
                            <hi rend="bold">Vandome, Robin</hi> (2013): “Introduction: Networks and
                        the Nineteenth-Century Periodical”, in: <hi rend="italic">American
                            Periodicals</hi> 23, 2: 93-104. </bibl>
                    <bibl>
                        <hi rend="bold">Ganascia, Jean G.</hi> / <hi rend="bold">Glaudes,
                            Pierre</hi> / <hi rend="bold">Del Lungo, Andrea</hi> (2014): “Automatic
                        Detection of Reuses and Citations in Literary Texts”, in: <hi rend="italic"
                            >Literary and Linguistic Computing</hi> 29, 3: 412-421. </bibl>
                    <bibl>
                        <hi rend="bold">Genette, Gérard</hi> (1997): <hi rend="italic">Palimpsests:
                            Literaure in the Second Degree</hi>. Lincoln, Neb: University of
                        Nebraska Press. </bibl>
                    <bibl>
                        <hi rend="bold">Haberman, Robb K.</hi> (2008): “Magazines, Presentation
                        Networks, and the Cultivation of Authorship in Post-Revolutionary America”,
                        in: <hi rend="italic">American Periodicals</hi> 18, 2: 141-162. </bibl>
                    <bibl>
                        <hi rend="italic bold">HaLebanon</hi> 1874-1883. </bibl>
                    <bibl>
                        <hi rend="italic bold">HaMagid</hi> 1874-1883. </bibl>
                    <bibl>
                        <hi rend="italic bold">HaMelitz</hi> 1874-1883. </bibl>
                    <bibl>
                        <hi rend="italic bold">HaTzfira</hi> 1874-1883. </bibl>
                    <bibl>
                        <hi rend="italic bold">Havazelet</hi> 1874-1883. </bibl>
                    <bibl>
                        <hi rend="bold">Kouts, Gideon</hi> (2013): <hi rend="italic">News and
                            History: Studies in History of the Hebrew and Jewish Press and
                            Communication</hi>. Jerusalem: The Zionist Library and Tel Aviv
                        University. [Hebrew]. </bibl>
                    <bibl>
                        <hi rend="bold">Lee, John</hi> (2007): “A Computational Model of Text Reuse
                        in Ancient Literary Texts”, in: Zaenen, Annie / van den Bosch, Antal (eds.):
                            <hi rend="italic">Proceedings of the 45th Annual Meeting of the
                            Association of Computational Linguistics</hi>. Prague: Association for
                        Computational Linguistics 472-479. </bibl>
                    <bibl>
                        <hi rend="bold">Lovepreet, Vishal G.</hi> / <hi rend="bold">Kumar,
                            Rohit</hi> (2020): “Survey on Plagiarism Detection Systems and Their
                        Comparison”, in: Behera, Himansu / Nayak, Janmenjoy / Naik, Bighnaraj /
                        Pelusi, Danilo (eds.): <hi rend="italic">Computational Intelligence in Data
                            Mining: Advances in Intelligent Systems and Computing</hi>. Singapore:
                        Springer 27-39. </bibl>
                    <bibl>
                        <hi rend="bold">Mintz, Alan</hi> (1995): “Introduction: The Many Rather Than
                        the One: On the Critical Study of Jewish Periodicals”, in: <hi rend="italic"
                            >Prooftexts</hi> 15, 1: 1-4. </bibl>
                    <bibl>
                        <hi rend="bold">Mosteller, Frederick</hi> / <hi rend="bold">Wallace, David
                            L.</hi> (1963): “Inference in an Authorship Problem: A Comparative Study
                        of Discrimination Methods Applied to the Authorship of the Disputed
                        Federalist Papers”, in: <hi rend="italic">Journal of the American
                            Statistical Association</hi> 58, 302: 275-309. </bibl>
                    <bibl>
                        <hi rend="bold">Murphy, J. Stephen</hi> (2014): “Introduction: Visualizing
                        Periodical Networks”, in: <hi rend="italic">The Journal of Modern Periodical
                            Studies</hi> 5, 1: iii-xv. </bibl>
                    <bibl>
                        <hi rend="bold">Olsen, Mark</hi> / <hi rend="bold">Horton, Russell</hi> /<hi
                            rend="bold"> Roe, Glenn</hi> (2010): “Something Borrowed: Sequence
                        Alignment and the Identification of Similar Passages in Large Text
                        Collections”, in: <hi rend="italic">Digital Studies / Le champ
                            numérique</hi> 2, 1 &lt;<ref
                            target="https://www.digitalstudies.org/articles/10.16995/dscn.258/"
                            >https://www.digitalstudies.org/articles/10.16995/dscn.258/</ref>&gt;. </bibl>
                    <bibl>
                        <hi rend="bold">Penslar, Derek</hi> (2000): “Introduction: The Press and the
                        Jewish Public Sphere”, in: <hi rend="italic">Jewish History</hi> 14: 3-7. </bibl>
                    <bibl>
                        <hi rend="bold">Segal, Zef</hi> (2019): “’A Letter to our Authors’: The
                        Problem of False News in Hatsfira, 1874”, in: <hi rend="italic">Kesher</hi>
                        52, 15-20. [Hebrew]. </bibl>
                    <bibl>
                        <hi rend="bold">Segal, Zef </hi>(forthcoming): “From a Local Periodical to a
                        Global Enterprise: Ha-Me’asef, 1896- 1914”, in: <hi rend="italic">Journal of
                            Historical Network Research</hi>. </bibl>
                    <bibl>
                        <hi rend="bold">Smith, David A.</hi> / <hi rend="bold">Cordell, Ryan</hi> /
                            <hi rend="bold">Dillon, Elizabeth M.</hi> (2013): “Infectious Texts:
                        Modeling Text Reuse in Nineteenth-Century Newspapers”, in: <hi rend="italic"
                            >2013 IEEE International Conference on Big Data</hi> 86-94. </bibl>
                    <bibl>
                        <hi rend="bold">Soffer, Oren</hi> (2007): <hi rend="italic">There Is No
                            Place for Pilpul! HaTzfira Journal and the Modernization of the
                            Socio-political Discourse</hi>. Jerusalem: Mossad Bialik Press with the
                        Center for Research on the History and Culture of Polish Jewry at the Hebrew
                        University. [Hebrew]. </bibl>
                    <bibl>
                        <hi rend="bold">Soffer, Oren</hi> (2009): “Why Hebrew? A Comparative
                        Analysis of Language Choice in the Early Hebrew Press”, in: <hi
                            rend="italic">Media History</hi> 15, 3: 253-269. </bibl>
                    <bibl>
                        <hi rend="bold">Stamatatos, Efstathios</hi> / <hi rend="bold">Fakotakis,
                            Nikos</hi> / <hi rend="bold">Kokkinakis, George</hi> (2000): “Text genre
                        detection using common word frequencies”, in: <hi rend="italic">COLING
                            2000</hi>. Proceedings of the 18th conference on Computational
                        linguistics 2: 808–814 DOI: https://doi.org/10.3115/992730.992763. </bibl>
                    <bibl>
                        <hi rend="bold">Unsworth, Anna</hi> / <hi rend="bold">Morton, Andrew Q.</hi>
                        (1981): “Mrs. Gaskell Anonymous: Some Unidentified Items in ‘Fraser's
                        Magazine’”, in: <hi rend="italic">Victorian Periodicals Review</hi> 14, 1:
                        24-31. </bibl>
                </listBibl>
            </div>
        </back>
    </text>
</TEI>
