?
Метод автоматической генерации модели управления глаголов русского языка
С. 227–235.
Klyshinskiy E., Kochetkova N. A.
In book
Т. 2. , Белгород: Белгородский государственный технологический университет им. В.Г. Шухова, 2012.
Derkacheva A., Sakirkina M., Kraev G. et al., /. 2026.
Comprehensive data on natural hazards and their consequences are crucial for effective for risk assessment, adaptation planning, and emergency response. However, many countries face challenges with fragmented, inconsistent, and inaccessible data, particularly regarding local-scale events. To address this data gap in Russia, we developed an end-to-end processing pipeline that scrapes news from various online sources, ...
Added: April 28, 2026
Lelik V., Eremicheva T., Morozova D. et al., В кн.: Когнитивная наука в Москве: новые исследования. Материалы конференции 21–22 июня 2023 г.: М.: «Буки Веди», Московский институт психоанализа, 2023. С. 274–279.
Some of the important conditions of the effectiveness of morphological analyzers are the correct recognition of unfamiliar words and successful morphological disambiguation. In this work, we evaluated the results of automatic processing of children’s spontaneous speech using the morphological analyzer MyStem. We analyzed the longitudinal spontaneous speech recordings of two bilingual children and their parents ...
Added: April 5, 2024
Северина Е. М., Ларионова М. Ч., Litera 2023 № 10 С. 211–222
The article considers a model of preparation of machine-readable (semantic) markup of texts for the Chekhov Digital project on the example of philological interpretation of individual significant elements of A. P. Chekhov's story "Death of an Official" and presentation of this information explicitly based on the standards of digital publication Text Encoding Initiative (TEI/XML). Based ...
Added: January 12, 2024
Kolmogorova A., Калинин А. А., В кн.: Язык и искусственный интеллект: Сборник статей по итогам конференции «Лингвистический форум 2020: Язык и искусственный интеллект».: Издательский дом ЯСК, 2023. С. 167–181.
In the paper, we discuss the problem of tools supposed to be effective for visualization of data achieved as result of running algorithms for emotional text analysis. We start by overviewing some technics used to visualize data in projects devoted to exploratory data analysis, sentiment-analysis and emotional text analysis. To continue, we suggest two variants ...
Added: October 31, 2023
Morkovkin A., Ilvovsky D., В кн.: ИТиС 2022: Сборник трудов 46-й междисциплинарной школы-конференции ИППИ РАН "Информационные технологии и системы 2022".: Институт проблем передачи информации им. А.А. Харкевича РАН, 2022.
The estimation of textual complexity is an important and relevant task in the field of natural language processing. For example, in the banking sector, according to experts, there is a trend towards increasing the complexity of texts in all areas of financial regulation, which makes them difficult to understand even by professionals. This can lead ...
Added: September 23, 2023
Bolshakova E. I., Sapin A., , in: Computational Linguistics and Intellectual Technologies: Papers from the Annual International Conference “Dialogue” (2021)Issue 20: Основной том.: -, 2021. P. 154–161.
Added: October 30, 2021
M.: Russian State University for the Humanitie, 2019.
The book includes 64 papers submitted to the International conference in computer linguistics and intellectual technologies Dialogue 2019 and presents a broad spectrum of theoretical and applied research of natural language description, language simulation, and creation of applied computer technologies. ...
Added: October 16, 2019
Лаврентьев А. М., Смирнов И. В., Соловьев Ф. Н. et al., Вопросы кибербезопасности 2019 № 4(32) С. 54–60
Цель исследования: разработка методики создания и автоматического анализа специальных корпусов текстов для последующего применения их в качестве обучающих выборок и определения дифференцирующих признаков в задачах классификации текстов.
Метод: применялись инструменты анализа корпусной платформы TXM, расширенной разработанными процедурами вычисления дополнительных характеристик текстов, таких как буквосочетания, псевдоосновы, именные группы, глагольные группы.
Полученные результаты: показано, что разработанные средства расширения ...
Added: August 10, 2019
Лаврентьев А. М., Соловьев Ф. Н., Chepovskiy A., В кн.: Труды международной конференции "Корпусная лингвистика - 2019".: СПб.: Издательство Санкт-Петербургского университета, 2019. С. 55–62.
Представлен опыт расширения возможностей платформы TXM за счет инструментов автоматической обработки текста (выделение псевдооснов, именных групп, анализ глагольного управления). В сочетании со стандартными функциями TXM (факторный анализ соответствий, специфичность и т.д.) они позволяют более эффективно осуществлять анализ специализированных корпусов, нацеленных, в частности, на выявление противоправного дискурса. ...
Added: July 8, 2019
Volkova L. L., В кн.: Новые информационные технологии в автоматизированных системах: материалы шестнадцатого научно-технического семинара.: М.: Московский государственный институт электроники и математики, 2013. С. 317–328.
В статье дан краткий обзор ключевых этапов развития машинной лингвистики в разрезе анализа и синтеза текста. Выделены проблемы работы с языком, являющиеся фундаментальными ограничениями, отделяющими существующий уровень развития отрасли от качественно нового. Рассмотрены перспективные теории, предлагающие новый подход к рассмотрению языка и открывающие возможность заглянуть за барьер машинной лингвистики. ...
Added: January 31, 2018
Durandin O. V., Strebkov D. Y., Hilal N. R., , in: Computational Linguistics and Intellectual Technologies: Proceedings of the Annual International Conference “Dialogue” (2016).: М.: Изд-во РГГУ, 2016. P. 1–13.
The paper presents work on automatic Arabic dialect classification and proposes machine learning classification method where training dataset consists of two corpora. The first one is a small corpus of manually dialectannotated instances. The second one contains big amount of instances that were grabbed from the Web automatically using word-marks—most unique and frequent dialectal words ...
Added: January 18, 2017
Kharlamov A. A., Ермоленко Т. В., Жонин А. А., В кн.: Открытые семантические технологии проектирования интеллектуальных систем = Open Semantic Technologies for Intelligent Systems (OSTIS-2014) : материалы IV междунар. науч.-техн. конф. (Минск, 20-22 февраля 2014 года).: Мн.: БГУИР, 2014. С. 161–168.
Описанный в статье подход к моделированию динамики процессов основан на технологии автоматического смыслового анализа текстовой информации. В процессе обработки текста формируется ассоциативная сеть, ключевые понятия которой, в том числе, лексические маркеры анализируемого процесса, ранжируются их смысловым весом. Взвешенный статусом маркера на шкале «хорошо - плохо», этот вес дает значение вклада маркера в характеристику состояния процесса. ...
Added: November 19, 2016
Kubatieva A., В кн.: I Молодежная международная конференция «Методы точных наук в востоковедении», 10-11 ноября 2015 г.: Материалы конференции.: СПб.: Издательство РХГА, 2015.
In this paper, we describe basic principles of POS-classifications and their modelling for POS-tagging of Chinese and statistical NLP systems. Using three available statistical POS-taggers, we conducted an experiment on POS-tagging of Chinese text to analyze quality evaluation, correspondence between POS-tags and categories assigned in different reference grammars. We also determine the basic rules of ...
Added: December 10, 2015
Malafeev A., , in: Computational Linguistics and Intellectual Technologies. Papers from the Annual International Conference “Dialogue” (2015)Issue 14(21).: M.: Russian State University for the Humanitie, 2015. P. 441–452.
Current trends in education, namely blended learning and computer-assisted language learning, underlie the growing interest to the task of automatically generating language exercises. Such automatic systems are especially in demand given the variability in language learning. Despite the abundance of resources for language learning, there is often a lack of specific exercises targeting a particular ...
Added: April 28, 2015
Klyshinskiy E., Kalachyov Y. B., Научная визуализация 2014 Т. 6 № 3 С. 96–104
В статье рассматривается метод визуального контроля полноты технической документации, изучаемой в ходе ее приемки. Полнота отчетной документации проверяется по тексту технического задания. Для построения визуального представления используются как цветные точечные диаграммы, так и статистическая информация о степени соответствия текстов, полученная с использованием предлагаемого метода. ...
Added: March 25, 2015
Dubov M., Mirkin B., Шаль А. А., Открытые системы. СУБД 2014 № 10 С. 15–17
Currently, automating of text processing and analysis is a main tendency of IT applications. As of this moment, there is no unified approach to the analysis and visualization of big volumes of text data. Our system LM Monitor (Latent Meaning Monitor) generates so-called reference graphs which can be considered part of the popular technology of ...
Added: December 16, 2014
Malafeev A., International Journal of Conceptual Structures and Smart Applications (IJCSSA) 2014 Vol. 2 No. 2 P. 20–35
This article presents an approach to the automatic generation of open cloze exercises based on arbitrary English text. The exercise format is similar to the open cloze test used in Cambridge English certificate exams (FCE, CAE, CPE). The presented method also makes it possible to adjust the difficulty of the resulting exercises to better suit ...
Added: November 29, 2014
Karpov N., International Journal of Advances in Computer Science and Its Applications 2014 Vol. 4 No. 2 P. 38–43
The algorithm to adapt lexical complexity in the news article which can be used as materials for learning language presented in the paper. We consider words substitution retrieval according to wordnet-based and corpus-based semantic relatedness. Two corpus-based similarity measures empirically tested: Vector Space Model and Distributional Semantic Model. This language processing algorithm has created as ...
Added: November 23, 2014
Malafeev A., В кн.: Homo Loquens: Актуальные вопросы лингвистики и методики преподавания иностранных языков (2014)Вып. 6.: СПб.: Отдел оперативной полиграфии НИУ ВШЭ – Санкт-Петербург, 2014. С. 372–381.
The article describes an effective method of automatic generation of word-formation exercises in the format similar to Cambridge certificate exams and Russian state exam. The input for the generation of exercises is English texts of arbitrary size and contents, which is especially relevant within the framework of teaching English for special or academic purposes. To ...
Added: October 15, 2014
Klyshinskiy E., Kalachyov Y. B., Zhadnov V. V., Научно-техническая информация. Серия 2: Информационные процессы и системы 2014 № 5 С. 11–15
Рассматривается новый метод автоматизации определения соответствия технического задания и итогового отчета в ходе его приемки. Предложенный метод позволяет экспертам получить предварительную оценку степени соответствия отчета техническому заданию. Используются выделение значимых фрагментов технического задания,поиск соответствующих им элементов отчета и проверка степени его покрытия. Разработанный метод,в отличие, например,от косинусной меры сходства, дает лучшее разделение отчетов по критерию ...
Added: June 30, 2014
Klyshinskiy E., Kalachyov Y. B., Zhadnov V. V., Информационные технологии в проектировании и производстве 2014 № 2 С. 68–72
The paper introduces a new method of technical reports verification by finding the correspondence between an agreed technical statement and a submitted report during its acceptance process. The method makes a preliminary evaluation of such correspondence. For these purposes we select significant fragments of technical statement, mark the coinciding parts of report and evaluate the ...
Added: June 3, 2014
М.: Институт прикладной математики им. М.В. Келдыша РАН, 2014.
Содержит материалы, представленные к рассмотрению на научно-практический семинар “Новые информационные технологии в автоматизированных системах”.
Представляет интерес для научных сотрудников, преподавателей, аспирантов и студентов, работающих по указанным научным направлениям. ...
Added: April 12, 2014