<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.3 20210610//EN" "JATS-journalpublishing1-3.dtd">
<article article-type="research-article" dtd-version="1.3" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xml:lang="ru"><front><journal-meta><journal-id journal-id-type="publisher-id">lingngu</journal-id><journal-title-group><journal-title xml:lang="ru">Вестник НГУ. Серия: Лингвистика и межкультурная коммуникация</journal-title><trans-title-group xml:lang="en"><trans-title>NSU Vestnik. Series: Linguistics and Intercultural Communication</trans-title></trans-title-group></journal-title-group><issn pub-type="ppub">1818-7935</issn><publisher><publisher-name>Новосибирский государственный университет</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.25205/1818-7935-2025-23-3-123-135</article-id><article-id custom-type="elpub" pub-id-type="custom">lingngu-1097</article-id><article-categories><subj-group subj-group-type="heading"><subject>Research Article</subject></subj-group><subj-group subj-group-type="section-heading" xml:lang="ru"><subject>КОМПЬЮТЕРНАЯ И ПРИКЛАДНАЯ ЛИНГВИСТИКА</subject></subj-group><subj-group subj-group-type="section-heading" xml:lang="en"><subject>COMPUTER AND APPLIED LINGUISTICS</subject></subj-group></article-categories><title-group><article-title>Методы повышения согласия аннотаторов при моделировании аргументационной структуры текста</article-title><trans-title-group xml:lang="en"><trans-title>Methods for Improving Inter-Annotator Agreement in Modelling Argumentation Structure of a Text</trans-title></trans-title-group></title-group><contrib-group><contrib contrib-type="author" corresp="yes"><contrib-id contrib-id-type="orcid">https://orcid.org/0000-0001-5946-9469</contrib-id><name-alternatives><name name-style="eastern" xml:lang="ru"><surname>Пименов</surname><given-names>И. С.</given-names></name><name name-style="western" xml:lang="en"><surname>Pimenov</surname><given-names>I. S.</given-names></name></name-alternatives><bio xml:lang="ru"><p>Пименов Иван Сергеевич, программист</p><p>Новосибирск</p></bio><bio xml:lang="en"><p>Ivan S. Pimenov, Coder</p><p>Novosibirsk</p></bio><email xlink:type="simple">pimenov.1330@yandex.ru</email><xref ref-type="aff" rid="aff-1"/></contrib><contrib contrib-type="author" corresp="yes"><contrib-id contrib-id-type="orcid">https://orcid.org/0000-0001-8412-9116</contrib-id><name-alternatives><name name-style="eastern" xml:lang="ru"><surname>Саломатина</surname><given-names>Н. В.</given-names></name><name name-style="western" xml:lang="en"><surname>Salomatina</surname><given-names>N. V.</given-names></name></name-alternatives><bio xml:lang="ru"><p>Саломатина Наталья Васильевна, кандидат физико-математических наук, старший научный сотрудник </p><p>Новосибирск</p></bio><bio xml:lang="en"><p>Natalia V. Salomatina, PhD, Senior Researcher</p><p>Novosibirsk</p></bio><email xlink:type="simple">salomatina_nv@live.ru</email><xref ref-type="aff" rid="aff-1"/></contrib></contrib-group><aff-alternatives id="aff-1"><aff xml:lang="ru"><institution>Институт систем информатики им. А. П. Ершова СО РАН</institution><country>Россия</country></aff><aff xml:lang="en"><institution>А. P. Ershov Institute of Informatics Systems</institution><country>Russian Federation</country></aff></aff-alternatives><pub-date pub-type="collection"><year>2025</year></pub-date><pub-date pub-type="epub"><day>11</day><month>02</month><year>2026</year></pub-date><volume>23</volume><issue>3</issue><fpage>123</fpage><lpage>135</lpage><permissions><copyright-statement>Copyright &amp;#x00A9; Пименов И.С., Саломатина Н.В., 2026</copyright-statement><copyright-year>2026</copyright-year><copyright-holder xml:lang="ru">Пименов И.С., Саломатина Н.В.</copyright-holder><copyright-holder xml:lang="en">Pimenov I.S., Salomatina N.V.</copyright-holder><license xml:lang="ru" license-type="creative-commons-attribution" xlink:href="https://creativecommons.org/licenses/by/4.0/" xlink:type="simple"><license-p>Данная работа распространяется под лицензией Creative Commons Attribution 4.0.</license-p></license><license xml:lang="en" license-type="creative-commons-attribution" xlink:href="https://creativecommons.org/licenses/by/4.0/" xlink:type="simple"><license-p>This work is licensed under a Creative Commons Attribution 4.0 License.</license-p></license></permissions><self-uri xlink:href="https://lingngu.elpub.ru/jour/article/view/1097">https://lingngu.elpub.ru/jour/article/view/1097</self-uri><abstract><p>Качество работы алгоритмов машинного обучения, применяемых для решения многих задач NLP, в том числе задачи автоматического извлечения структур аргументации из текста, существенно зависит от единообразия разметки корпуса данных. В разметке объемных корпусов, необходимых для обучения алгоритмов, как правило, участвуют несколько аннотаторов, что снижает ее единообразие или согласованность. В данной работе с целью улучшения согласованности предложены методы модификации аргументативной разметки, построенной разметчиками в виде графов аргументации с помощью инструментов платформы ArgNetBank Studio согласно стандарту Argument Interchange Format. Графы содержат вершины с текстом утверждений (посылок, заключений), связанные в аргумент ребрами через вершины-схемы (типы аргументов из списка Уолтона). В работе предлагаются способы автоматической модификации графов для случаев расхождений между разметчиками в определении аргументативности высказываний, относящихся к «периферии» рассуждений, и отдельных типов аргументов. Повышение согласованности на уровне высказываний осуществляется двумя путями: 1) удалением листовых вершин графа, представленных лишь в одной разметке и 2) преобразованием цепочки из двух аргументов в один при выполнении определенных условий. Рост согласованности типов аргументов осуществляется за счет замены узкоспециализированных схем на более общие, а также по правилам исходя из принадлежности схем к функциональным группам, содержащим близкие по семантике и по специфике реализации в рассуждениях. Количественной оценкой согласованности обычно служат коэффициенты согласия. Подсчитанный коэффициент α Криппендорфа показал значительный рост согласия в разметке после модификации корпуса из 160 графов аргументации 80 коротких научных статей (на 25 % для аргументативных высказываний и на 19 % для схем). Однако оценки качества распознавания классификатором MLP предложений с аргументацией продемонстрировали только 5%-й рост F-меры. Четыре исследуемые схемы рассуждения CausetoEffect, VerbalClassification, Example, PracticalReasoning на оригинальном и модифицированном корпусе показали еще меньший прирост F-меры: не более 2 %. Сделан вывод о недостаточности только прироста коэффициента согласия для существенного улучшения оценок качества распознавания.</p></abstract><trans-abstract xml:lang="en"><p>The efficiency of machine learning algorithms applied to various NLP tasks, particularly the automatic extraction of argumentation structures from texts, heavily depends on datasets annotation consistency. As a rule, annotation of large corpora for algorithms training relies on a joint effort of several annotators, detrimental to its consistency. The present article describes methods of modifying argumentation annotations for improving their consistency. Annotations in the study take form of argumentation graphs constructed in accordance with the Argument Interchange Format standard, using ArgNetBank Studio tools. These graphs contain nodes with statements text (for premises, conclusions) aggregated into arguments with edges through scheme nodes (which correspond to argument types from Walton’s compendium). The study proposes methods for automatic modification of graphs in case of annotators’ disagreement in identifying argumentative statements and arguments of specific types. At the level of statements, consistency improvement relies on two procedures: 1) removal of leaf nodes present only in one of the graphs; 2) transformation of two-argument sequences into one-argument upon fulfillment of specific conditions. For argument types, agreement increases through replacing infrequent narrow-focused schemes with more common general ones, as well as by applying a hierarchical system of schemes substitution rules based on their functional classification (a functional group unites schemes similar in semantic and textual expression properties). A typical quantitative way of measuring annotation consistency is employing inter-annotator agreement coefficients. Calculation of the Krippendorff α agreement coefficient demonstrates a considerable increase of consistency upon modifying a corpus of 160 annotations for 80 short scientific articles in Russian: the increase equals 25 % for argumentative statements and 19 % for argument types. However, an experiment in identifying argumentative sentences with an MLP classifier shows only a 5 % increase of F-measure after modifying the corpus. The increase of F-measure for identifying four argumentation types under analysis (Cause to Effect, Verbal Classification, Example, Practical Reasoning) is even less: no more than 2 %. We arrive at a conclusion that an improvement of the inter-annotator agreement coefficient is by itself insufficient for a considerable increase in identification efficiency values in practice.</p></trans-abstract><kwd-group xml:lang="ru"><kwd>корпусная лингвистика</kwd><kwd>научные статьи</kwd><kwd>разметка аргументации</kwd><kwd>коэффициенты согласия</kwd><kwd>Argument Mining</kwd></kwd-group><kwd-group xml:lang="en"><kwd>corpus linguistics</kwd><kwd>scientific articles</kwd><kwd>argumentation annotation</kwd><kwd>inter-annotator agreement</kwd><kwd>argument mining</kwd></kwd-group><funding-group><funding-statement xml:lang="ru">Исследование выполнено за счет гранта Российского научного фонда № 23-21-00325, https://rscf.ru/ project/23-21-00325/</funding-statement><funding-statement xml:lang="en">The study was supported by the Russian Science Foundation grant No. 23-21-00325, https://rscf.ru/project/23-21-00325/</funding-statement></funding-group></article-meta></front><back><ref-list><title>References</title><ref id="cit1"><label>1</label><citation-alternatives><mixed-citation xml:lang="ru">Пименов И. С. Сочетаемость аргументов разных функциональных групп в научных текстах // Филологические науки. Вопросы теории и практики. Тамбов: Грамота, 2022. № 11. С. 3672–3680.</mixed-citation><mixed-citation xml:lang="en">Castro S. Fast Krippendorff: Fast computation of Krippendorff’s alpha agreement measure, 2017. https://github.com/pln-fing-udelar/fast-krippendorff.</mixed-citation></citation-alternatives></ref><ref id="cit2"><label>2</label><citation-alternatives><mixed-citation xml:lang="ru">Пименов И. С. Анализ расхождений в аргументационной разметке научных статей на русском языке // Вестник НГУ. Серия: Лингвистика и межкультурная коммуникация. 2023. Т. 21, № 2. С. 89–104. DOI 10.25205/1818-7935-2023-21-2-89-104.</mixed-citation><mixed-citation xml:lang="en">Fishcheva I., Kotelnikov E. Cross-Lingual Argumentation Mining for Russian Texts. Proc. Of the 8th International Conference “Analysis of Images, Social Networks and Texts” (Kazan), 2019, pp. 134–144.</mixed-citation></citation-alternatives></ref><ref id="cit3"><label>3</label><citation-alternatives><mixed-citation xml:lang="ru">Сидорова Е.А., Ахмадеева И.Р., Загорулько Ю.А., Серый А.С., Шестаков В.К. Платформа для исследования аргументации в научно-популярном дискурсе // Онтология проектирования. 2020. Т. 10, № 4(38). С. 489–502.</mixed-citation><mixed-citation xml:lang="en">Fishcheva I., Goloviznina V., and Kotelnikov E. Traditional machine learning and deep learning models for argumentation mining in russian texts. Computational Linguistics and Intellectual Technologies: Proceedings of the International Conference “Dialog-2021”, 2021, pp. 246–258.</mixed-citation></citation-alternatives></ref><ref id="cit4"><label>4</label><citation-alternatives><mixed-citation xml:lang="ru">Castro S. Fast Krippendorff: Fast computation of Krippendorff’s alpha agreement measure, 2017. https://github.com/pln-fing-udelar/fast-krippendorff.</mixed-citation><mixed-citation xml:lang="en">Kotelnikov E., Loukachevitch N., Nikishina I., Panchenko A. RuArg-2022: argument mining evaluation. Proceedings of the International Conference “Dialogue 2022”, 2022, pp. 1–16.</mixed-citation></citation-alternatives></ref><ref id="cit5"><label>5</label><citation-alternatives><mixed-citation xml:lang="ru">Fishcheva I., Kotelnikov E. Cross-Lingual Argumentation Mining for Russian Texts // Proc. Of the 8th International Conference “Analysis of Images, Social Networks and Texts” (Kazan), 2019, pp. 134–144.</mixed-citation><mixed-citation xml:lang="en">Krippendorff K. Content analysis: An introduction to its methodology, 3rd edition. Thousand Oaks, CA: Sage, 2013.</mixed-citation></citation-alternatives></ref><ref id="cit6"><label>6</label><citation-alternatives><mixed-citation xml:lang="ru">Fishcheva I., Goloviznina V., and Kotelnikov E. Traditional machine learning and deep learning models for argumentation mining in russian texts // Computational Linguistics and Intellectual Technologies: Proceedings of the International Conference “Dialog-2021”. 2021. P. 246–258.</mixed-citation><mixed-citation xml:lang="en">Pimenov I. S. Compatibility of Arguments from Different Functional Groups in Scientific Texts. Philology. Theory &amp; Practice. Tambov, Gramota, 2022, vol. 11. pp. 3672−3680. (in Russ.)</mixed-citation></citation-alternatives></ref><ref id="cit7"><label>7</label><citation-alternatives><mixed-citation xml:lang="ru">Kotelnikov E., Loukachevitch N., Nikishina I., Panchenko A. RuArg-2022: argument mining evaluation // Proceedings of the International Conference “Dialogue 2022”. 2022. P. 1–16.</mixed-citation><mixed-citation xml:lang="en">Pimenov I. S. Analyzing Disagreements in Argumentation Annotation of Scientific Texts in Russian Language. Vestnik NSU. Series: Linguistics and Intercultural Communications, 2023, vol. 21, no. 2, pp. 89–104. (in Russ.) DOI 10.25205/1818-7935-2023-21-2-89-104.</mixed-citation></citation-alternatives></ref><ref id="cit8"><label>8</label><citation-alternatives><mixed-citation xml:lang="ru">Krippendorff K. Content analysis: An introduction to its methodology, 3rd edition. Thousand Oaks, CA: Sage, 2013.</mixed-citation><mixed-citation xml:lang="en">Pimenov I. S., Salomatina N. V. An Automatic Method for Standartizing Argumentative Annotations across Annotators. Proc. 2024 IEEE 25th International Conference of Young Professionals in Electron Devices and Materials (EDM), 28 June 2024 – 02 July 2024. DOI: 10.1109/EDM61683.2024.10615176.</mixed-citation></citation-alternatives></ref><ref id="cit9"><label>9</label><citation-alternatives><mixed-citation xml:lang="ru">Pimenov I. S., Salomatina N. V. An Automatic Method for Standartizing Argumentative Annotations across Annotators // Proc. 2024 IEEE 25th International Conference of Young Professionals in Electron Devices and Materials (EDM), 28 June 2024 – 02 July 2024. DOI: 10.1109/EDM61683.2024.10615176</mixed-citation><mixed-citation xml:lang="en">Rahwan I., Reed C. The argument interchange format. Argumentation in artificial intelligence / ed. Rahwan I. and Simari G., Springer, 2009, pp. 383–402.</mixed-citation></citation-alternatives></ref><ref id="cit10"><label>10</label><citation-alternatives><mixed-citation xml:lang="ru">Rahwan I., Reed C. The argument interchange format. Argumentation in artificial intelligence // ed. Rahwan I. and Simari G., Springer, 2009, pp. 383–402.</mixed-citation><mixed-citation xml:lang="en">Sidorova Е. А., Akhmadeeva I. R., Zagorulko Yu. A., Sery A. S., Shestakov V. K. Research platform for the study of argumentation in popular science discourse. Ontology design, 2020, vol. 10, no. 4 (38), pp. 489–502. (in Russ.)</mixed-citation></citation-alternatives></ref><ref id="cit11"><label>11</label><citation-alternatives><mixed-citation xml:lang="ru">Skeppstedt M., Peldzus A., Stede M. More or less controlled elicitation of argumentative text: Enlarging a microtext corpus via crowdsourcing // Proceedings of the 5th Workshop on Argument Mining, Brussels, Belgium, 2018, pp. 155–163.</mixed-citation><mixed-citation xml:lang="en">Skeppstedt M., Peldzus A., Stede M. More or less controlled elicitation of argumentative text: Enlarging a microtext corpus via crowdsourcing. Proceedings of the 5th Workshop on Argument Mining, Brussels, Belgium, 2018, pp. 155–163.</mixed-citation></citation-alternatives></ref><ref id="cit12"><label>12</label><citation-alternatives><mixed-citation xml:lang="ru">Stab C. and Gurevych I. Annotating argument components and relations in persuasive essays // Proceedings of COLING 2014а, the 25th International Conference on Computational Linguistics: Technical Papers. 2014a. P. 1501–1510.</mixed-citation><mixed-citation xml:lang="en">Stab C. and Gurevych I. Annotating argument components and relations in persuasive essays. Proceedings of COLING 2014а, the 25th International Conference on Computational Linguistics: Technical Papers, 2014a, pp. 1501–1510.</mixed-citation></citation-alternatives></ref><ref id="cit13"><label>13</label><citation-alternatives><mixed-citation xml:lang="ru">Stab C., Gurevych I. Identifying argumentative discourse structures in persuasive essays // Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP). 2014b. P. 46–56.</mixed-citation><mixed-citation xml:lang="en">Stab C. and Gurevych I. Identifying argumentative discourse structures in persuasive essays. Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP), 2014b, pp. 46–56.</mixed-citation></citation-alternatives></ref><ref id="cit14"><label>14</label><citation-alternatives><mixed-citation xml:lang="ru">Stab C., Gurevych I. Parsing Argumentation Structures in Persuasive Essays // Computational Linguistics. 2017. Vol. 43, no. 3. P. 619−659.</mixed-citation><mixed-citation xml:lang="en">Stab C. and Gurevych I. Parsing Argumentation Structures in Persuasive Essays. Computational Linguistics, 2017, vol. 43, no. 3, pp. 619–659.</mixed-citation></citation-alternatives></ref><ref id="cit15"><label>15</label><citation-alternatives><mixed-citation xml:lang="ru">Teruel M., Cardellino C., Cardellino F., Alemany L., Villata S. Increasing Argument Annotation Reproducibility by Using Inter-annotator Agreement to Improve Guidelines // Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018). Miyazaki, Japan, 2018.</mixed-citation><mixed-citation xml:lang="en">Teruel M., Cardellino C., Cardellino F., Alemany L., Villata S. Increasing Argument Annotation Reproducibility by Using Inter-annotator Agreement to Improve Guidelines. Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018). Miyazaki, Japan, 2018.</mixed-citation></citation-alternatives></ref><ref id="cit16"><label>16</label><citation-alternatives><mixed-citation xml:lang="ru">Walton D., Reed C., Macagno F. Argumentation schemes Fundamentals of critical argumentation. New York: Cambridge University Press, 2008, 443 p.</mixed-citation><mixed-citation xml:lang="en">Walton D., Reed C., Macagno F. Argumentation schemes Fundamentals of critical argumentation. New York, Cambridge University Press, 2008, 443 p.</mixed-citation></citation-alternatives></ref></ref-list><fn-group><fn fn-type="conflict"><p>The authors declare that there are no conflicts of interest present.</p></fn></fn-group></back></article>
