<?xml version="1.0" encoding="UTF-8"?>
<TEI xmlns="http://www.tei-c.org/ns/1.0">
    <teiHeader>
        <fileDesc>
            <titleStmt>
                <title>User-centered design of a query editor for bio-bibliographical data retrieval by domain experts</title>
                <author>
                    <persName>
                        <surname>Philipp</surname>
                        <forename>Luisa</forename>
                    </persName>
                    <affiliation>Beuth Hochschule für Technik, Germany</affiliation>
                    <email>luisa.philipp@beuth-hochschule.de</email>
                </author>
                <author>
                    <persName>
                        <surname>Kreutel</surname>
                        <forename>Jörn</forename>
                    </persName>
                    <affiliation>Beuth Hochschule für Technik, Germany</affiliation>
                    <email>joern.kreutel@beuth-hochschule.de</email>
                </author>
            </titleStmt>
            <editionStmt>
                <edition>
                    <date>2021-05-27T13:17:00Z</date>
                </edition>
            </editionStmt>
            <publicationStmt>
                <publisher>Elisabeth Burr, University of Leipzig</publisher>
                <address>
                    <addrLine>Beethovenstr. 15</addrLine>
                    <addrLine>04107 Leipzig</addrLine>
                    <addrLine>Germany</addrLine>
                    <addrLine>Elisabeth Burr</addrLine>
                </address>
            </publicationStmt>
            <sourceDesc>
                <p>Converted from a Word document</p>
            </sourceDesc>
        </fileDesc>
        <encodingDesc>
            <appInfo>
                <application ident="DHCONVALIDATOR" version="1.22">
                    <label>DHConvalidator</label>
                </application>
            </appInfo>
        </encodingDesc>
        <profileDesc>
            <textClass>
                <keywords scheme="ConfTool" n="category">
                    <term>Paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="subcategory">
                    <term>Short Paper</term>
                </keywords>
                <keywords scheme="ConfTool" n="keywords">
                    <term>User-centered Design</term>
                    <term>Querying</term>
                    <term>Data Retrieval</term>
                </keywords>
                <keywords scheme="ConfTool" n="topics">
                    <term>Designing</term>
                    <term>Web development</term>
                    <term>Network Analysis</term>
                    <term>Spatial Analysis</term>
                    <term>Visualization</term>
                    <term>Collaboration</term>
                    <term>Publishing</term>
                    <term>Sharing</term>
                    <term>Meta: Assessing</term>
                    <term>Meta: CommunityBuilding</term>
                    <term>Meta: ProjectManagement</term>
                    <term>BibliographicListings</term>
                    <term>Infrastructure</term>
                    <term>Persons</term>
                    <term>Tools</term>
                    <term>Software</term>
                    <term>Literature</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>not applicable</term>
                    <term>English</term>
                </keywords>
            </textClass>
        </profileDesc>
    </teiHeader>
    <text>
        <body>
            <p>In current practice of software development for digital humanities (DH) applications,
                the consideration of an application’s usability as a major factor for user
                acceptance can still not be taken for granted (Thoden et al. 2017), (Lehenmeier /
                Burghardt 2019). However, particularly for projects that focus on data collection in
                some field, not only the initial project success, but also the longer-term
                dissemination of project results as well as continuous scientific cooperation based
                on a project’s outcome will depend on the accessibility of data collections to
                domain experts beyond the actual project participants. This lets usability appear to
                be a relevant curation aspect with regard to the collected data.<note place="foot"
                    xml:id="ftn1" n="1">
                    <p rend="footnote text"> See the overview in Zuiderwijk et al. (2020), which
                        mentions poor usability of “complex user interfaces” for querying as an
                        inhibitor of data reuse. <lb/>
                    </p>
                </note> Yet, considering recent publications, neither the question of how domain
                experts may intuitively access structured data using appropriate query tools as a
                prerequisite of various types of analyses, nor the usability of such tools seem to
                receive major attention in DH, where existing work has a strong focus on the SPARQL
                language (Harris / Seaborne 2013), rather than dealing with the issue of query
                creation in a more generic way.<note place="foot" xml:id="ftn2" n="2">
                    <p rend="footnote text"> See Heibi et al. (2017) for a recent proposal of a
                        generic query infrastructure based on SPARQL that includes a usability
                        analysis, as well as Russel / Smart (2008), Ambrus et al. (2010) and Soylu
                        et al. (2016), which are not focussing on DH scenarios, though. For the
                        relevance of structured querying for domain experts, see, e.g., a two-week
                        summer school course from 2018 on “Asking questions to data in the
                        humanities: right, correct, efficient” dealing with XQuery, SQL and SPARQL
                            (<ref target="https://esu.culintec.de/?q=node/942"
                            >https://esu.culintec.de/?q=node/942</ref>, last accessed on May 20,
                        2021). As for the EADH2021 conference, however, the fine-grained
                        classification scheme for contributions via ConfTool, in contrast to the
                        CfP, does not mention querying as an activity of its own in the context of
                        data analysis. <lb/>
                    </p>
                </note>
            </p>
            <p>Our contribution will report on ongoing development of a web-based software platform
                for collecting and analysing bio-bibliographical data for literary studies on GDR
                    authors,<note place="foot" xml:id="ftn3" n="3">
                    <p rend="footnote text"> See <ref target="http://www.ddr-literatur.de/"
                            >http://www.ddr-literatur.de/</ref> for further details. The project is
                        funded by Deutsche Forschungsgemeinschaft (DFG), grant number 419244741. The
                        authors would like to thank their project partners from Humboldt University
                        of Berlin and the Berlin Brandenburg Academy of Sciences and Humanities, who
                        participated in the user interviews underlying the work presented here.
                        <lb/>
                    </p>
                </note> and will present findings from the evaluation of a query editor that allows
                domain experts to compose complex queries. Applying a user-centered design
                methodology (Abras et al. 2004) and given the information needs of the project
                participants, the editor was supposed to provide an “expert mode” alternative to an
                existing free text search interface and should particularly support the formulation
                of complex queries involving relations between various entities of the underlying
                data model and constraints on the latter.<note place="foot" xml:id="ftn4" n="4">
                    <p> See, e.g., the following example queries: “ <hi rend="italic">What is the
                            gender/year-of-birth/birthplace etc. of authors associated with a given
                            institution (e.g., a writer’s association, university, etc.)?</hi> ” – “
                            <hi rend="italic">Which jobs did writers have before/after their first
                            publication?</hi> ” – “ <hi rend="italic">Over time, which were the
                            publishing houses issuing publications of authors that were associated
                            with a given institution?</hi>
                    </p>
                </note> Technically, the editor should not depend on a particular data storage
                paradigm or query language, but should employ an abstract query representation
                mappable onto particular query languages like JPQL (DeMichiel / Jungmann 2017) or
                SPARQL. In order to be portable to alternative domains and as the domain model for
                data collection had been iteratively developed by domain experts and software
                engineers, it was expected to adapt to the respective model structure. </p>
            <p>Starting with interviews for finding out the German literature domain experts’
                information needs, overall technical affinity and experience with query interfaces,
                and considering their design preferences known from cooperation within the project,
                the design of the editor basically adheres to a form-based, rather than an
                icon-based or diagram-based approach (see Catarci et al. 1997; Lloret-Gazo 2016 for
                this distinction). As the figure below shows, it uses embedded structures for
                expressing conditions on entities that involve relations to other entities,<note
                    place="foot" xml:id="ftn5" n="5">
                    <p rend="footnote text"> The underlying abstract query representation allows the
                        usage of variables for coindexation of entities and is, hence, more
                        expressive than a mere tree structure. Integration and end user tests of
                        this advanced feature will be done at a later stage of the project,
                        though.</p>
                </note> where relations can be both expressed on the basis of an entity’s own
                outgoing associations and based on their “incoming” usage within other entities,
                thus allowing flexible entry points into query formulation. Natural language style
                labels (e.g., “<hi rend="italic">Gib mir alle</hi>”, <hi
                    rend="italic">“mit... Attribut</hi>”, etc.) support intuitive readability of the
                created query. </p>
            <figure>
                <graphic n="1001" width="15.980833333333333cm" height="10.6045cm" url="Pictures/85d842e1a3ea17a17f33add24c41b3d6.png" rend="inline"/>
            </figure>
            <p>Representation of a query retrieving birth events for female persons who have regularly completed an education in Berlin. In a similar way, the projection of the query, which specifies which attributes of the selected entities shall actually be included in the query result, can be formulated. For this example, the projection could include, e.g., the names of the persons, the dates and places of birth and the types and periods of the retrieved education events.</p>
            <p>Following Jakob Nielsen’s 5-users-rule for usability testing (Nielsen 2000)<note
                    place="foot" xml:id="ftn6" n="6">
                    <p rend="footnote text"> See, e.g., Faulkner (2003) for an investigation of the
                        limitations of this principle.</p>
                </note> , the editor has, so far, been evaluated with six domain experts, whose
                majority was familiar with the domain model in terms of entity types and their
                attributes and associations as they had already applied that model for data entry.
                Each tester carried out five supervised retrieval tasks that were expressed in
                written natural language, where complexity gradually increased. Applying the
                think-aloud method (Van Someren et al. 1994), observable obstacles and explicit
                questions of test persons were tracked by the supervisor. All tasks could be
                completed successfully by all testers, where the majority of unclear issues could be
                solved by the testers themselves without requiring help by the supervisor. Where
                help was necessary, this was mainly due to conceptual aspects of the domain model
                that had not shown up during the latter’s usage for data collection, like the notion
                of an entity’s “incoming” associations. However, once the respective questions had
                been clarified, testers were able to successfully apply their increased knowledge in
                subsequent tasks. </p>
            <p>Given the objective to make the platform available to external domain experts, the
                findings from the tests show, on the one hand, that future versions of the editor
                should improve access for users unfamiliar with the domain model.<note place="foot"
                    xml:id="ftn7" n="7">
                    <p rend="footnote text"> This could be achieved, e.g., by interactive tutorials,
                        which could apply proposals for knowledge transfer based on “example-driven
                        modelling” (Bak et al. 2013) to the case of query formulation.</p>
                </note> On the other hand, provided the latter knowledge, the tests also reveal that
                particular technical experience or familiarity with “expert mode” search interfaces
                is not a prerequisite for successful creation of even complex queries that involve
                entity relations. Consistently with this observation of the testers’ objective
                achievements, also the evaluation of the users’ subjective experience on the basis
                of the <hi rend="italic">user experience questionnaire</hi> (Laugwitz et al. 2006)
                shows an overall positive attitude and perception of, among other aspects, the
                editor’s learnability and ease of use. Hence, the evaluation appears promising with
                respect to the platform’s objective to support domain experts’ research interests by
                providing user interfaces with good usability. </p>
        </body>
        <back>
            <div type="bibliogr">
                <listBibl>
                    <head>Bibliography</head>
                    <bibl>
                        <hi rend="bold">Abras, Chadia</hi> / <hi rend="bold">Maloney-Krichmar,
                            Diane</hi> / <hi rend="bold">Preece, Jenny</hi> (2004): "User-centered
                        design", in: Bainbridge, William (ed.): <hi rend="italic">Encyclopedia of
                            Human-Computer Interaction</hi>. Thousand Oaks: Sage Publications
                        445-456. </bibl>
                    <bibl>
                        <hi rend="bold">Ambrus, Oszkár</hi> / <hi rend="bold">Möller, Knud</hi> /
                            <hi rend="bold">Handschuh, Siegfried</hi> (2010): <hi rend="italic"
                            >Konduit VQB: a visual query builder for SPARQL on the social semantic
                            desktop</hi>. Workshop on Visual Interfaces to the Social and Semantic
                        Web (VISSW2010), IUI2010, Feb 7, 2010, Hong Kong, China &lt;<ref
                            target="http://ceur-ws.org/Vol-565/paper4.pdf"
                            >http://ceur-ws.org/Vol-565/paper4.pdf</ref>&gt; [02.09.2021].</bibl>
                    <bibl>
                        <hi rend="bold">Bak, Kacper</hi> / <hi rend="bold">Zayan, Dina</hi> / <hi
                            rend="bold">Czarnecki, Krzysztof</hi> / <hi rend="bold">Antkiewicz,
                            Michał</hi> / <hi rend="bold">Diskin, Zinovy</hi> / <hi rend="bold"
                            >Wasowski, Andrzej</hi> / <hi rend="bold">Rayside, Derek</hi> (2013):
                        "Example-driven modeling: model=abstractions+examples", in: <hi
                            rend="italic">35th International Conference on Software Engineering
                            (ICSE)</hi>, IEEE 1273-1276. </bibl>
                    <bibl>
                        <hi rend="bold">Catarci, Tiziana</hi> / <hi rend="bold">Costabilr</hi> / <hi
                            rend="bold">Maria Francesca</hi> / <hi rend="bold">Levialdi,
                            Stefano</hi> / <hi rend="bold">Batini, Carlo</hi> (April 1997): "Visual
                        query systems for databases: A survey", in: <hi rend="italic">Journal of
                            Visual Languages &amp; Computing</hi> 8, 2: 215-260. </bibl>
                    <bibl>
                        <hi rend="bold">DeMichiel, Linda</hi> / <hi rend="bold">Jungmann, Lukas</hi>
                        (2017): <hi rend="italic">JSR 338: Java Persistence API</hi>, Version 2.2.
                        Technical report, Oracle Corporation. </bibl>
                    <bibl>
                        <hi rend="bold">Faulkner, Laura</hi> (August 2003): "Beyond the five-user
                        assumption: Benefits of increased sample sizes in usability testing", in:
                            <hi rend="italic">Behavior Research Methods, Instruments, &amp;
                            Computers</hi> 35, 3: 379-383. </bibl>
                    <bibl>
                        <hi rend="bold">Harris, Steve</hi> / <hi rend="bold">Seaborne, Andy</hi>
                        (2013): <hi rend="italic">SPARQL 1.1 Query Language</hi> &lt;<ref
                            target="https://www.w3.org/TR/sparql11-query/"
                            >https://www.w3.org/TR/sparql11-query/</ref>&gt; [01.05.2021]. </bibl>
                    <bibl>
                        <hi rend="bold">Heibi, Ivan</hi> / <hi rend="bold">Peroni, Silvio</hi> / <hi
                            rend="bold">Shotton, David</hi> (2017): "OSCAR: a customisable tool for
                        free-text search over SPARQL endpoints", in: González-Beltrán, Alejandra /
                        Osborne, Francesco / Peroni, Silvio / Vahdati, Sahar (eds.): <hi
                            rend="italic">Semantics, Analytics, Visualization</hi>. 3rd
                        International Workshop, SAVE-SD 2017, Perth, Australia, April 3, 2017, and
                        4th International Workshop, SAVE-SD 2018, Lyon, France, April 24, 2018,
                        Revised Selected Papers. Berlin / Heidelberg: Springer 121-137 </bibl>
                    <bibl>
                        <hi rend="bold">Laugwitz, Bettina</hi> / <hi rend="bold">Schrepp,
                            Martin</hi> / <hi rend="bold">Held, Theo</hi> (2006): "Konstruktion
                        eines Fragebogens zur Messung der User Experience von Softwareprodukten",
                        in: Heinecke, Andreas M. / Paul, Hansjürgen (eds.): <hi rend="italic">Mensch
                            und Computer 2006: Mensch und Computer im Strukturwandel</hi>. München /
                        Wien: Oldenbourg 125-134. </bibl>
                    <bibl>
                        <hi rend="bold">Lehenmeier, Constantin</hi> / <hi rend="bold">Burghardt,
                            Manuel</hi> (2019): "Usability statt Frustration", in: Draude, Claude /
                        Lange, Martin, / Sick, Bernhard (eds.): <hi rend="italic">INFORMATIK 2019
                            Workshops</hi>. Lecture Notes in Informatics (LNI). Bonn: Gesellschaft
                        für Informatik Bonn 2019 97-106. </bibl>
                    <bibl>
                        <hi rend="bold">Lloret-Gazo, Jorge</hi> (2016): "A survey on visual query
                        systems in the web era", in: Hartmann, Sven / Ma, Hui (eds.): <hi
                            rend="italic">International Conference on Database and Expert Systems
                            Applications</hi>. 27th International Conference, DEXA 2016, Porto,
                        Portugal, September 5-8, 2016, Proceedings, Part II. Berlin / Heidelberg:
                        Springer 343-351. </bibl>
                    <bibl>
                        <hi rend="bold">Nielsen, Jakob</hi> (March 2000): <hi rend="italic">Why you
                            only need to test with 5 users</hi> &lt;<ref
                            target="https://www.nngroup.com/articles/why-you-only-need-to-test-with-5-users/"
                            >https://www.nngroup.com/articles/why-you-only-need-to-test-with-5-users/</ref>&gt;
                        [01.05.2021]. </bibl>
                    <bibl>
                        <hi rend="bold">Russell, Alistair</hi> / <hi rend="bold">Smart, Paul R.</hi>
                        (2008): "NITELIGHT: A graphical editor for SPARQL queries", in: <hi
                            rend="italic">Proceedings of ISWC (Posters and Demos)</hi> 401 &lt;<ref
                            target="http://ceur-ws.org/Vol-401/iswc2008pd_submission_11.pdf"
                            >http://ceur-ws.org/Vol-401/iswc2008pd_submission_11.pdf</ref>&gt;
                        [02.09.2021].</bibl>
                    <bibl>
                        <hi rend="bold">Soylu, Ahmet</hi> / <hi rend="bold">Giese, Martin</hi> / <hi
                            rend="bold">Jiménez-Ruiz, Ernesto</hi> / <hi rend="bold">Kharlamov,
                            Evgeny</hi> / <hi rend="bold">Zheleznyakov, Dmitriy</hi> / <hi
                            rend="bold">Horrocks, Ian</hi> (April 2016): "Ontology-based end-user
                        visual query formulation: Why, what, who, how, and which?", in: <hi
                            rend="italic">Universal Access in the Information Society </hi>16:
                        435–467. </bibl>
                    <bibl>
                        <hi rend="bold">Thoden, Klaus</hi> / <hi rend="bold">Stiller, Juliane</hi> /
                            <hi rend="bold">Bulatovic, Natasa</hi> / <hi rend="bold">Meiners,
                            Hanna-Lena</hi> / <hi rend="bold">Boukhelifa, Nadia</hi> (April 2017):
                        "User-centered design practices in digital humanities - experiences from
                        DARIAH and CENDARI", in: <hi rend="italic">ABI Technik</hi> 37: 2-11. </bibl>
                    <bibl>
                        <hi rend="bold">Van Someren, Maarten</hi> / <hi rend="bold">Barnard,
                            Yvonne</hi> / <hi rend="bold">Sandberg, Jacobijn</hi> (1994): <hi
                            rend="italic">The think aloud method: a practical approach to modelling
                            cognitive processes</hi>. London: AcademicPress </bibl>
                    <bibl>
                        <hi rend="bold">Zuiderwijk, Anneke</hi> / <hi rend="bold">Shinde,
                            Rhythima</hi> / <hi rend="bold">Jeng, Wei</hi> (2020): "What drives and
                        inhibits researchers to share and use open research data? A systematic
                        literature review to analyze factors influencing open research data
                        adoption", in: <hi rend="italic">PloS one</hi> 15, 9: e0239283. </bibl>
                </listBibl>
            </div>
        </back>
    </text>
</TEI>
