<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.3 20210610//EN" "JATS-journalpublishing1-3.dtd">
<article article-type="research-article" dtd-version="1.3" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xml:lang="ru"><front><journal-meta><journal-id journal-id-type="publisher-id">intechngu</journal-id><journal-title-group><journal-title xml:lang="ru">Вестник НГУ. Серия: Информационные технологии</journal-title><trans-title-group xml:lang="en"><trans-title>Vestnik NSU. Series: Information Technologies</trans-title></trans-title-group></journal-title-group><issn pub-type="ppub">1818-7900</issn><issn pub-type="epub">2410-0420</issn><publisher><publisher-name>НГУ</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.25205/1818-7900-2021-19-1-61-79</article-id><article-id custom-type="elpub" pub-id-type="custom">intechngu-155</article-id><article-categories><subj-group subj-group-type="heading"><subject>Research Article</subject></subj-group><subj-group subj-group-type="section-heading" xml:lang="ru"><subject>Статьи</subject></subj-group></article-categories><title-group><article-title>Статистический метод количественной оценки понятности иностранных славянских языков для русскоязычного читателя</article-title><trans-title-group xml:lang="en"><trans-title>Quantitative Estimation of Intelligibility of Foreign Slavic Languages: Case of Russian Native Speakers</trans-title></trans-title-group></title-group><contrib-group><contrib contrib-type="author" corresp="yes"><contrib-id contrib-id-type="orcid">https://orcid.org/0000-0002-4020-488X</contrib-id><name-alternatives><name name-style="eastern" xml:lang="ru"><surname>Клышинский</surname><given-names>Э. С.</given-names></name><name name-style="western" xml:lang="en"><surname>Klyshinsky</surname><given-names>E. S.</given-names></name></name-alternatives><email xlink:type="simple">klyshinsky@mail.ru</email><xref ref-type="aff" rid="aff-1"/></contrib></contrib-group><aff-alternatives id="aff-1"><aff xml:lang="ru"><institution>Институт прикладной математики им. М. В. Келдыша РАН</institution><country>Россия</country></aff><aff xml:lang="en"><institution>Keldysh Institute of Applied Mathematics RAS</institution><country>Russian Federation</country></aff></aff-alternatives><pub-date pub-type="collection"><year>2021</year></pub-date><pub-date pub-type="epub"><day>24</day><month>05</month><year>2021</year></pub-date><volume>19</volume><issue>1</issue><fpage>61</fpage><lpage>79</lpage><permissions><copyright-statement>Copyright &amp;#x00A9; Клышинский Э.С., 2021</copyright-statement><copyright-year>2021</copyright-year><copyright-holder xml:lang="ru">Клышинский Э.С.</copyright-holder><copyright-holder xml:lang="en">Klyshinsky E.S.</copyright-holder><license xml:lang="ru" license-type="creative-commons-attribution" xlink:href="https://creativecommons.org/licenses/by/4.0/" xlink:type="simple"><license-p>Данная работа распространяется под лицензией Creative Commons Attribution 4.0.</license-p></license><license xml:lang="en" license-type="creative-commons-attribution" xlink:href="https://creativecommons.org/licenses/by/4.0/" xlink:type="simple"><license-p>This work is licensed under a Creative Commons Attribution 4.0 License.</license-p></license></permissions><self-uri xlink:href="https://intechngu.elpub.ru/jour/article/view/155">https://intechngu.elpub.ru/jour/article/view/155</self-uri><abstract><p>Разбирается вопрос понимаемости иностранного текста на славянском языке для неподготовленного информанта. Целью статьи было выяснить, какую долю слов иностранного текста информанты смогут понять при условии, что они не знакомы с этим языком. Для определения понимаемости текста мы использовали параллельный текст с пропущенными словами. В русской версии текста пропускалась часть слов, задача информанта - восстановить эти слова, используя в качестве подсказки параллельный текст на одном из славянских языков: украинском, белорусском, польском, чешском, словацком, сербском, словенском и болгарском. Часть информантов использовалась в качестве контрольной группы, и параллельный текст им не предъявлялся. Мы высказали гипотезу о том, что понятность текста на иностранном языке может быть определена как увеличение доли корректно восстанавливаемых слов в группе, которой предъявляется параллельный текст на иностранном языке, над долей слов, корректно восстановленных контрольной группой. Результаты экспериментов подтвердили нашу гипотезу. Также мы разделили все пары «пропущенное слово - перевод» на четыре группы: полные и частичные когнаты, генетические когнаты, не когнаты и ложные друзья. Корреляция средней понятности текста по всем информантам для данного языка с долей полных и частичных когнатов составила 0.7, тогда как для остальных групп была отрицательной. За счет этого можно утверждать, что понятность иностранного текста по большей части определяется долей полных когнатов, но при этом зависит от некоторых других параметров. Результаты экспериментов и программное обеспечение для их анализа размещены по адресу https:// github.com/klyshinsky/mutual_intelligibility_Russian.</p></abstract><trans-abstract xml:lang="en"><p>In this article, we investigate the issue of intelligibility of a foreign Slavic text for a Russian-speaking person which don’t know this language. The aim of this article is to find out what is the percentage of intelligible words in foreign text for such a person. As a main measuring tool, we used parallel cloze tests with omitted words in the Russian part. The task was to restore omitted words using the foreign part of a test (written in Ukrainian, Belorussian, Polish, Czech, Slovak, Serbian, Slovene, and Bulgarian languages) as a clue. As a baseline, we used a control group which solved a test without the foreign part. Our hypothesis was that the foreign text intelligibility could be defined as a difference between the mean percentage of correctly restored words for a group used a parallel text and the same percentage for a control group. The results of our experiments proved our hypothesis. All the pairs “omitted word - its translation” was divided into four groups: full and partial cognates, genetic cognates, non-cognates and false friends. The correlation between the mean intelligibility of a text in a given foreign language and the percentage of full and partial cognates was as high as 0.7; the same correlation for the other word groups was negative but not so deep. Therefore, we can state that the foreign text intelligibility is defined by the percentage of full and partial cognates but that is not the only parameter. The gathered data, containing the used tests, users’ answers and their background, and the software for its analysis is placed at https://github.com/klyshinsky/mutual_intelligibility_Russian.</p></trans-abstract><kwd-group xml:lang="ru"><kwd>понятность текста</kwd><kwd>славянские языки</kwd><kwd>тест с пропусками</kwd><kwd>корреляция</kwd><kwd>квантитативная лингвистика</kwd></kwd-group><kwd-group xml:lang="en"><kwd>text intelligibility</kwd><kwd>Slavic languages</kwd><kwd>cloze test</kwd><kwd>correlation</kwd><kwd>quantitative linguistics</kwd></kwd-group></article-meta></front><back><ref-list><title>References</title><ref id="cit1"><label>1</label><citation-alternatives><mixed-citation xml:lang="ru">Бейкер М. Атомы языка: грамматика в темном поле сознания. М.: Изд-во ЛКИ, 2008. 272 с.</mixed-citation><mixed-citation xml:lang="en">Бейкер М. Атомы языка: грамматика в темном поле сознания. М.: Изд-во ЛКИ, 2008. 272 с.</mixed-citation></citation-alternatives></ref><ref id="cit2"><label>2</label><citation-alternatives><mixed-citation xml:lang="ru">Moberg, J., Gooskens C., Nerbonne J., Vaillette N. Conditional Entropy Measures Intelligibility among Related Languages. Lot Occasional Series, 2007, vol. 7, p. 51-66.</mixed-citation><mixed-citation xml:lang="en">Moberg, J., Gooskens C., Nerbonne J., Vaillette N. Conditional Entropy Measures Intelligibility among Related Languages. Lot Occasional Series, 2007, vol. 7, p. 51-66.</mixed-citation></citation-alternatives></ref><ref id="cit3"><label>3</label><citation-alternatives><mixed-citation xml:lang="ru">Gooskens C. The Contribution of Linguistic Factors to the Intelligibility of Closely Related Languages. Journal of Multilingual and Multicultural Development, 2007, vol. 28, no. 6, p. 445-467.</mixed-citation><mixed-citation xml:lang="en">Gooskens C. The Contribution of Linguistic Factors to the Intelligibility of Closely Related Languages. Journal of Multilingual and Multicultural Development, 2007, vol. 28, no. 6, p. 445-467.</mixed-citation></citation-alternatives></ref><ref id="cit4"><label>4</label><citation-alternatives><mixed-citation xml:lang="ru">Golubović J., Gooskens C. Mutual intelligibility between West and South Slavic languages. Russ Linguist, 2015, no. 39, p. 351-373.</mixed-citation><mixed-citation xml:lang="en">Golubović J., Gooskens C. Mutual intelligibility between West and South Slavic languages. Russ Linguist, 2015, no. 39, p. 351-373.</mixed-citation></citation-alternatives></ref><ref id="cit5"><label>5</label><citation-alternatives><mixed-citation xml:lang="ru">Gooskens C., Swarte F. Linguistic and extra-linguistic predictors of mutual intelligibility between Germanic languages. Nordic Journal of Linguistics, 2017. no. 40 (2), p. 123-147. DOI 10.1017/S0332586517000099</mixed-citation><mixed-citation xml:lang="en">Gooskens C., Swarte F. Linguistic and extra-linguistic predictors of mutual intelligibility between Germanic languages. Nordic Journal of Linguistics, 2017. no. 40 (2), p. 123-147. DOI 10.1017/S0332586517000099</mixed-citation></citation-alternatives></ref><ref id="cit6"><label>6</label><citation-alternatives><mixed-citation xml:lang="ru">Kyjánek L., Haviger J. The Measurement of Mutual Intelligibility between West-Slavic Languages. Journal of Quantitative Linguistics, 2019, vol. 26, iss. 3, p. 205-230. DOI 10.1080/ 09296174.2018.1464546</mixed-citation><mixed-citation xml:lang="en">Kyjánek L., Haviger J. The Measurement of Mutual Intelligibility between West-Slavic Languages. Journal of Quantitative Linguistics, 2019, vol. 26, iss. 3, p. 205-230. DOI 10.1080/ 09296174.2018.1464546</mixed-citation></citation-alternatives></ref><ref id="cit7"><label>7</label><citation-alternatives><mixed-citation xml:lang="ru">Keatley C. W. History of bilingualism research in cognitive psychology. In: Harris R. J. (ed.). Cognitive processing in bilinguals. Elsevier, 1992, p. 15-49.</mixed-citation><mixed-citation xml:lang="en">Keatley C. W. History of bilingualism research in cognitive psychology. In: Harris R. J. (ed.). Cognitive processing in bilinguals. Elsevier, 1992, p. 15-49.</mixed-citation></citation-alternatives></ref><ref id="cit8"><label>8</label><citation-alternatives><mixed-citation xml:lang="ru">Grainger J. Visual word recognition in bilinguals. In: Schreuder R., Weltens B. (eds.). The bilingual lexicon. Amsterdam, 1993, p. 11-26.</mixed-citation><mixed-citation xml:lang="en">Grainger J. Visual word recognition in bilinguals. In: Schreuder R., Weltens B. (eds.). The bilingual lexicon. Amsterdam, 1993, p. 11-26.</mixed-citation></citation-alternatives></ref><ref id="cit9"><label>9</label><citation-alternatives><mixed-citation xml:lang="ru">Lemhöfer K., Dijkstra T. Recognizing cognates and interlingual homographs: Effects of code similarity in language-specific and generalized lexical decision. Memory &amp; Cognition, 2004, no. 32 (4), p. 533-550. DOI 10.3758/BF03195845</mixed-citation><mixed-citation xml:lang="en">Lemhöfer K., Dijkstra T. Recognizing cognates and interlingual homographs: Effects of code similarity in language-specific and generalized lexical decision. Memory &amp; Cognition, 2004, no. 32 (4), p. 533-550. DOI 10.3758/BF03195845</mixed-citation></citation-alternatives></ref><ref id="cit10"><label>10</label><citation-alternatives><mixed-citation xml:lang="ru">Hammarström H. Counting Languages in Dialect Continua Using the Criterion of Mutual Intelligibility. Journal of Quantitative Linguistics, 2008, vol. 15, no. 1, p. 34-45.</mixed-citation><mixed-citation xml:lang="en">Hammarström H. Counting Languages in Dialect Continua Using the Criterion of Mutual Intelligibility. Journal of Quantitative Linguistics, 2008, vol. 15, no. 1, p. 34-45.</mixed-citation></citation-alternatives></ref><ref id="cit11"><label>11</label><citation-alternatives><mixed-citation xml:lang="ru">Коряков Ю. Б. Проблема «язык или диалект» и попытка лексикостатистического подхода // Вопросы языкознания. 2017. № 6. C. 79-101. DOI 10.31857/S0373658X0003839-1</mixed-citation><mixed-citation xml:lang="en">Коряков Ю. Б. Проблема «язык или диалект» и попытка лексикостатистического подхода // Вопросы языкознания. 2017. № 6. C. 79-101. DOI 10.31857/S0373658X0003839-1</mixed-citation></citation-alternatives></ref><ref id="cit12"><label>12</label><citation-alternatives><mixed-citation xml:lang="ru">Heeringa W. J. Measuring Dialect Pronunciation Differences using Levenshtein Distance. PhD thesisю Groningen, 2004, 315 p.</mixed-citation><mixed-citation xml:lang="en">Heeringa W. J. Measuring Dialect Pronunciation Differences using Levenshtein Distance. PhD thesisю Groningen, 2004, 315 p.</mixed-citation></citation-alternatives></ref><ref id="cit13"><label>13</label><citation-alternatives><mixed-citation xml:lang="ru">Клышинский Э. С., Логачева В. К., Белобокова Ю. А. Понимаемость текста на иностранном языке: случай славянских языков // Препринты ИПМ им. М. В. Келдыша. 2017. № 13. 23 с. DOI 10.20948/prepr-2017-13</mixed-citation><mixed-citation xml:lang="en">Клышинский Э. С., Логачева В. К., Белобокова Ю. А. Понимаемость текста на иностранном языке: случай славянских языков // Препринты ИПМ им. М. В. Келдыша. 2017. № 13. 23 с. DOI 10.20948/prepr-2017-13</mixed-citation></citation-alternatives></ref><ref id="cit14"><label>14</label><citation-alternatives><mixed-citation xml:lang="ru">Zarei A. A., Ab M. A. The Contribution of Word Formation, Code Mixing, Multiple Choice, and Gap Filling Tasks to L2 Vocabulary Comprehension and Production. International Journal of Language Learning and Applied Linguistics World, 2013, no. 4 (1), p. 7-55.</mixed-citation><mixed-citation xml:lang="en">Zarei A. A., Ab M. A. The Contribution of Word Formation, Code Mixing, Multiple Choice, and Gap Filling Tasks to L2 Vocabulary Comprehension and Production. International Journal of Language Learning and Applied Linguistics World, 2013, no. 4 (1), p. 7-55.</mixed-citation></citation-alternatives></ref><ref id="cit15"><label>15</label><citation-alternatives><mixed-citation xml:lang="ru">Ackerman P. L., Beier M. E., Bowen K. R. Explorations of crystallized intelligence: Completion tests, cloze tests, and knowledge. Learning and Individual Differences, 2000, vol. 12, iss. 1, p. 105-121.</mixed-citation><mixed-citation xml:lang="en">Ackerman P. L., Beier M. E., Bowen K. R. Explorations of crystallized intelligence: Completion tests, cloze tests, and knowledge. Learning and Individual Differences, 2000, vol. 12, iss. 1, p. 105-121.</mixed-citation></citation-alternatives></ref><ref id="cit16"><label>16</label><citation-alternatives><mixed-citation xml:lang="ru">Ageeva E., Tyers F. M., Forcada M. L., Pérez-Ortiz J. A. Evaluating machine translation for assimilation via a gaplling task. In: EAMT-2015: 18th Annual Conference of the European Association for Machine Translation, 2015, p. 137-144.</mixed-citation><mixed-citation xml:lang="en">Ageeva E., Tyers F. M., Forcada M. L., Pérez-Ortiz J. A. Evaluating machine translation for assimilation via a gaplling task. In: EAMT-2015: 18th Annual Conference of the European Association for Machine Translation, 2015, p. 137-144.</mixed-citation></citation-alternatives></ref><ref id="cit17"><label>17</label><citation-alternatives><mixed-citation xml:lang="ru">Ягунова Е. В. Исследование избыточности русского звучащего текста // Тр. Ин-та лингвистических исследований. СПб.: Наука, 2010. Т. 4, ч. 2. С. 90-114.</mixed-citation><mixed-citation xml:lang="en">Ягунова Е. В. Исследование избыточности русского звучащего текста // Тр. Ин-та лингвистических исследований. СПб.: Наука, 2010. Т. 4, ч. 2. С. 90-114.</mixed-citation></citation-alternatives></ref><ref id="cit18"><label>18</label><citation-alternatives><mixed-citation xml:lang="ru">Ferrer i Cancho R. The variation of Zipf’s law in human language. The European Physical Journal B - Condensed Matter and Complex Systems, 2005, no. 44, p. 249-257.</mixed-citation><mixed-citation xml:lang="en">Ferrer i Cancho R. The variation of Zipf’s law in human language. The European Physical Journal B - Condensed Matter and Complex Systems, 2005, no. 44, p. 249-257.</mixed-citation></citation-alternatives></ref><ref id="cit19"><label>19</label><citation-alternatives><mixed-citation xml:lang="ru">Кочеткова Н. А., Клышинский Э. С., Ермаков П. Д. Подчиняются ли составные конструкции закону Ципфа? // Системный администратор. 2016. № 11. С. 89-95.</mixed-citation><mixed-citation xml:lang="en">Кочеткова Н. А., Клышинский Э. С., Ермаков П. Д. Подчиняются ли составные конструкции закону Ципфа? // Системный администратор. 2016. № 11. С. 89-95.</mixed-citation></citation-alternatives></ref><ref id="cit20"><label>20</label><citation-alternatives><mixed-citation xml:lang="ru">Bouckaert R., Lemey P., Dunn M. et al. Mapping the Origins and Expansion of the Indo-European Language Family. Science, 2012, vol. 337, p. 957-960.</mixed-citation><mixed-citation xml:lang="en">Bouckaert R., Lemey P., Dunn M. et al. Mapping the Origins and Expansion of the Indo-European Language Family. Science, 2012, vol. 337, p. 957-960.</mixed-citation></citation-alternatives></ref><ref id="cit21"><label>21</label><citation-alternatives><mixed-citation xml:lang="ru">Klyshinskiy E., Karpik O. V. Quantitative Evaluation of Syntax Similarity. Mathematica Montisnigri, 2019, vol. 46, p. 123-132. DOI 10.20948/mathmon-2019-46-11</mixed-citation><mixed-citation xml:lang="en">Klyshinskiy E., Karpik O. V. Quantitative Evaluation of Syntax Similarity. Mathematica Montisnigri, 2019, vol. 46, p. 123-132. DOI 10.20948/mathmon-2019-46-11</mixed-citation></citation-alternatives></ref><ref id="cit22"><label>22</label><citation-alternatives><mixed-citation xml:lang="ru">Lopukhina A., Lopukhin K., Nosyrev G. Automated word sense frequency estimation for Russian nouns. In: Lyashevskaya O., Kopotev M., Mustajoki A. (eds.). Quantitative approaches to the Russian language. Routledge, 2018, p. 79-94. DOI 10.4324/9781315105048-4</mixed-citation><mixed-citation xml:lang="en">Lopukhina A., Lopukhin K., Nosyrev G. Automated word sense frequency estimation for Russian nouns. In: Lyashevskaya O., Kopotev M., Mustajoki A. (eds.). Quantitative approaches to the Russian language. Routledge, 2018, p. 79-94. DOI 10.4324/9781315105048-4</mixed-citation></citation-alternatives></ref></ref-list><fn-group><fn fn-type="conflict"><p>The authors declare that there are no conflicts of interest present.</p></fn></fn-group></back></article>
