<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//TaxonX//DTD Taxonomic Treatment Publishing DTD v0 20100105//EN" "../../nlm/tax-treatment-NS0.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:tp="http://www.plazi.org/taxpub" article-type="research-article" dtd-version="3.0" xml:lang="en">
  <front>
    <journal-meta>
      <journal-id journal-id-type="publisher-id">109</journal-id>
      <journal-id journal-id-type="index">urn:lsid:arphahub.com:pub:3dc5f44e-8666-58db-bc76-a455210e8891</journal-id>
      <journal-title-group>
        <journal-title xml:lang="en">JUCS - Journal of Universal Computer Science</journal-title>
        <abbrev-journal-title xml:lang="en">jucs</abbrev-journal-title>
      </journal-title-group>
      <issn pub-type="ppub">0948-695X</issn>
      <issn pub-type="epub">0948-6968</issn>
      <publisher>
        <publisher-name>Journal of Universal Computer Science</publisher-name>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.3897/jucs.86745</article-id>
      <article-id pub-id-type="publisher-id">86745</article-id>
      <article-categories>
        <subj-group subj-group-type="heading">
          <subject>Research Article</subject>
        </subj-group>
        <subj-group subj-group-type="scientific_subject">
          <subject>I.7 - DOCUMENT AND TEXT PROCESSING</subject>
          <subject>K.3 - COMPUTERS AND EDUCATION</subject>
          <subject>Topic M - Knowledge Management</subject>
        </subj-group>
      </article-categories>
      <title-group>
        <article-title>The evaluation of a semi-automatic authoring tool for knowledge extraction in the AC&amp;NL Tutor</article-title>
      </title-group>
      <contrib-group content-type="authors">
        <contrib contrib-type="author" corresp="yes">
          <name name-style="western">
            <surname>Grubišić</surname>
            <given-names>Ani</given-names>
          </name>
          <email xlink:type="simple">ani@pmfst.hr</email>
          <uri content-type="orcid">https://orcid.org/0000-0003-4313-7851</uri>
          <xref ref-type="aff" rid="A1">1</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Stankov</surname>
            <given-names>Slavomir</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0001-8997-7050</uri>
          <xref ref-type="aff" rid="A2">2</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Žitko</surname>
            <given-names>Branko</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0001-8946-0916</uri>
          <xref ref-type="aff" rid="A1">1</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Šarić-Grgić</surname>
            <given-names>Ines</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0002-9247-8890</uri>
          <xref ref-type="aff" rid="A1">1</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Gašpar</surname>
            <given-names>Angelina</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0002-4472-8648</uri>
          <xref ref-type="aff" rid="A1">1</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Brajković</surname>
            <given-names>Emil</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0002-1726-2029</uri>
          <xref ref-type="aff" rid="A3">3</xref>
        </contrib>
        <contrib contrib-type="author" corresp="no">
          <name name-style="western">
            <surname>Vasić</surname>
            <given-names>Daniel</given-names>
          </name>
          <uri content-type="orcid">https://orcid.org/0000-0002-7713-8396</uri>
          <xref ref-type="aff" rid="A3">3</xref>
        </contrib>
      </contrib-group>
      <aff id="A1">
        <label>1</label>
        <addr-line content-type="verbatim">University of Split, Split, Croatia</addr-line>
        <institution>University of Split</institution>
        <addr-line content-type="city">Split</addr-line>
        <country>Croatia</country>
      </aff>
      <aff id="A2">
        <label>2</label>
        <addr-line content-type="verbatim">Unaffiliated, Split, Croatia</addr-line>
        <institution>Unaffiliated</institution>
        <addr-line content-type="city">Split</addr-line>
        <country>Croatia</country>
      </aff>
      <aff id="A3">
        <label>3</label>
        <addr-line content-type="verbatim">University of Mostar, Mostar, Bosnia and Herzegovina</addr-line>
        <institution>University of Mostar</institution>
        <addr-line content-type="city">Mostar</addr-line>
        <country>Bosnia and Herzegovina</country>
      </aff>
      <author-notes>
        <fn fn-type="corresp">
          <p>Corresponding author: Ani Grubišić (<email xlink:type="simple">ani@pmfst.hr</email>).</p>
        </fn>
        <fn fn-type="edited-by">
          <p>Academic editor: </p>
        </fn>
      </author-notes>
      <pub-date pub-type="collection">
        <year>2023</year>
      </pub-date>
      <pub-date pub-type="epub">
        <day>28</day>
        <month>08</month>
        <year>2023</year>
      </pub-date>
      <volume>29</volume>
      <issue>8</issue>
      <fpage>866</fpage>
      <lpage>891</lpage>
      <uri content-type="arpha" xlink:href="http://openbiodiv.net/99DCF4C1-A23E-51DE-93C3-F43C53924679">99DCF4C1-A23E-51DE-93C3-F43C53924679</uri>
      <history>
        <date date-type="received">
          <day>19</day>
          <month>05</month>
          <year>2022</year>
        </date>
        <date date-type="accepted">
          <day>25</day>
          <month>04</month>
          <year>2023</year>
        </date>
      </history>
      <permissions>
        <copyright-statement>Ani Grubišić, Slavomir Stankov, Branko Žitko, Ines Šarić-Grgić, Angelina Gašpar, Emil Brajković, Daniel Vasić</copyright-statement>
        <license license-type="creative-commons-attribution" xlink:href="https://creativecommons.org/licenses/by-nd/4.0/" xlink:type="simple">
          <license-p>This is an open access article distributed under the terms of the Creative Commons Attribution License (CC BY-ND 4.0). This license allows reusers to copy and distribute the material in any medium or format in unadapted form only, and only so long as attribution is given to the creator. The license allows for commercial use.</license-p>
        </license>
      </permissions>
      <abstract>
        <label>Abstract</label>
        <p>This paper describes and evaluates the performance of a semi-automatic authoring tool (SAAT) for knowledge extraction in the AC&amp;NL Tutor, highlighting its strengths and weaknesses. We assessed the accuracy of automatic annotation tasks (Part-of-Speech tagging, Name Entity Recognition, Dependency parsing, and Coreference Resolution) performed on a dataset of 160 sentences from unstructured Wikipedia text on a computer. We compared the automatic annotations to the gold standard, created after human post-editing and validation. Human-error analysis included 3769 words, 582 subsentences, 1129 questions, 917 propositions, 1020 concepts, and 667 relations. It resulted in the error type classification and the set of custom rules further used for automatic error identification and correction. The results showed that an average of 68.7% of the error corrections referred to CoreNLP performance and 31.3% to the SAAT extraction algorithms. Our main contributions include an integrated approach to the comprehensive pre-processing of the text, knowledge extraction and visualization; the consolidated evaluation of natural language processing tasks and knowledge extraction output (sentences, subsentences, questions, concept maps) and the newly developed reference dataset. </p>
      </abstract>
      <funding-group>
        <award-group>
          <funding-source>
            <named-content content-type="funder_name">Office of Naval Research</named-content>
            <named-content content-type="funder_identifier">100000006</named-content>
            <named-content content-type="funder_doi">http://doi.org/10.13039/100000006</named-content>
          </funding-source>
        </award-group>
      </funding-group>
    </article-meta>
  </front>
</article>
