?
Да-да-да-да-да! И другие конструкции повседневной русской речи, состоящие из контактно повторяющихся слов, в контексте исследования просодических моделей языка
С. 162–175.
In book
Институт русского языка им. В.В. Виноградова РАН, 2018.
Kolmogorova A., Явшиц Е. В., Сугян А. Х. et al., Вестник Новосибирского государственного университета. Серия: Лингвистика и межкультурная коммуникация 2026 Т. 24 № 1 С. 113–127
Технологии автоматического распознавания речи (АРР) к настоящему моменту стали базовым компонентом бэкенд-инфраструктуры различных голосовых и когнитивных ассистентов, обеспечивая доступ пользователей к коммерческим сервисам и государственным платформам. Однако из-за неравномерного распределения данных, на которых модели АРР обучаются, острой становится проблема предвзятости моделей при работе с атипичной речью. Цель исследования заключается в оценке качества распознавания и сравнении ...
Added: August 28, 2026
Diachkova M., Lelik V., Dorofeeva S. et al., Language Resources and Evaluation 2026 Vol. 60 Article 68
This article presents the Russian Language-Monolingual corpus (RusLan-M, v.1.0), a longitudinal multimedia collection of early child speech from two Russian-speaking monolingual children: Tosya (ages 0;10–3;10, 246 recordings) and Yasha (ages 1;04–3;00, 42 recordings). The corpus consists of approximately 41 h (2,454 min.) of video recordings and 35,386 child utterances, available with transcriptions in the CHAT ...
Added: August 17, 2026
Britov I., Tyumeneva E., Glazunova S., , in: Ngoại giao, biên phiên dịch và hợp tác quốc tế của Việt Nam trong kỷ nguyên mới.: Hà Nội: Nhà xuất bản Đại học quốc gia Hà Nội, 2026. P. 417–427.
The article analyzes the Russian experience in writing textbooks on translation from Vietnamese into Russian and from Russian into Vietnamese. On the basis of five textbooks written in Russia over the past thirty years, forms and methods of teaching both oral and written translation are revealed. A comparative description of each of the presented textbooks ...
Added: August 11, 2026
Kharlamova D. S., Journal of the European Second Language Association 2026 Vol. 10 No. 1 P. 1–16
Formulaic language may help language learners in second language (L2) acquisition. However, interference with the first language (L1) can also cause errors in L2 production. The present paper explores the possibilities of detecting L1 Russian interference errors connected with phraseologisms in English learner texts with a fine-tuned Transformer-based neural network. Across a dataset of 3,600 ...
Added: August 6, 2026
Kharlamova D. S., Русский язык в научном освещении 2024 № 2 (48) С. 129–149
В статье исследуется история возникновения и эволюции идиомы от корки до корки. Мы использовали словарные данные и материалы НКРЯ и Google Books, чтобы установить примерное время появления данного выражения, а также предложить возможные объяснения его происхождения. Развитие конструкции от корки до корки связано с расширением ее сочетаемости: на этом пути можно выделить несколько этапов, которые ...
Added: August 6, 2026
Zubov V., Вопросы лексикографии 2026 № 40 С. 64–86
The article addresses the problem of selecting and systematizing data for the study of pronunciation variation in contemporary Russian and proposes a solution in the form of a specialized database of codified equivalent pronunciation variants (e.g., simmétriya / simmetríya “symmetry”). The article presents a methodology for identifying, selecting, and organizing such variants into a database. ...
Added: July 23, 2026
М.: Институт русского языка им. В.В. Виноградова РАН, 2026.
Сборник тезисов Пятнадцатых Шмелёвских чтений (К 100-летию со дня рождения академика Дмитрия Николаевича Шмелева) Жизнь слова: Научное наследие академика Д. Н. Шмелева в контексте современности. Охватывает разные аспекты современной русистики: от исторической лексикологии до современных трансформаций прагматики и семантики слов. ...
Added: June 23, 2026
Skripka N., Russian linguistics 2026 Vol. 50 Article 11
This paper presents the first in-depth corpus-based study of a previously overlooked syntactic variation in Russian: the competition between juxtapositional (Nominative) and possessive-like (Genitive) encoding of the second noun (the term) in specificational constructions (e.g., ponjatie čest’ (notion.NOM honor.NOM) vs. ponjatie česti (notion.NOMhonor.GEN) ‘the notion of honor’). While typological research has established cross-linguistic preferences for one encoding strategy over another, intralinguistic variation ...
Added: May 18, 2026
Zubov V., Осадчая М. А., Риехакайнен Е. И., Вестник Санкт-Петербургского университета. Язык и литература 2026 Т. 23 № 1 С. 99–119
The article is a part of a comprehensive study of the linguistic characteristics of teacher’s speech, which contribute to the success of the pedagogical discourse. Based on a survey of secondary school students and an analysis of previous research in the field, non-syntactic pauses of hesitation were chosen as the object of the study, i. ...
Added: April 29, 2026
Zubov V., Elena Riekhakaynen, , in: Proceedings of the Workshop on Cognitive Aspects of the Lexicon @ LREC-COLING 2024.: European Language Resources Association (ELRA), 2024. P. 129–132.
Variability is one of the important features of natural speech and a challenge for spoken word recognition models and automatic speech recognition systems. We conducted two preliminary experiments aimed at finding out whether native Russian speakers regard differently certain types of pronunciation variation when the variants are equally possible according to orthoepic norms. In the ...
Added: April 19, 2026
Глазкова А. В., Смаль И. В., Lyashevskaya O. et al., Доклады Российской академии наук. Математика, информатика, процессы управления (ранее - Доклады Академии Наук. Математика) 2025 Т. 527 С. 146–155
This paper presents a study on the effectiveness of discriminative methods for abbreviation lemmatization in Russian texts. Unlike generative approaches, discriminative models select the optimal lemma from a fixed set of candidates, eliminating the risk of generating grammatically incorrect word forms. For the first time in Russian language processing, we conduct a comprehensive analysis of ...
Added: March 10, 2026
Afanasev I., Glazkova A., Lyashevskaya O. et al., , in: Proceedings of the 10th Workshop on Slavic Natural Language Processing (Slavic NLP 2025).: Association for Computational Linguistics, 2025. P. 157–170.
Pre-trained language models have significantly advanced natural language processing (NLP), particularly in analyzing languages with complex morphological structures. This study addresses lemmatization for the Russian language, the errors in which can critically affect the performance of information retrieval, question answering, and other tasks. We present the results of experiments on generative lemmatization using pre-trained language ...
Added: March 10, 2026
Glazkova A., Lyashevskaya O., Morozov D. et al., Journal of Mathematical Sciences 2025 Vol. 546 P. 32–47
This paper addresses the task of lemmatizing abbreviations in the Russian language. Abbreviation lemmatization is particularly challenging, as it involves not only transforming a word into its normal form but also correctly expanding the abbreviation. We explore two approaches to this task, both leveraging large pretrained language models. The first approach is generative, where the ...
Added: March 10, 2026