• A
  • A
  • A
  • АБВ
  • АБВ
  • АБВ
  • A
  • A
  • A
  • A
  • A
Обычная версия сайта
  • RU
  • EN
  • HSE University
  • Publications
  • Book chapter
  • The Use of Khislavichi Lect Morphological Tagging to Determine its Position in the East Slavic Group
  • RU
  • EN
Расширенный поиск
Высшая школа экономики
Национальный исследовательский университет
Priority areas
  • business informatics
  • economics
  • engineering science
  • humanitarian
  • IT and mathematics
  • law
  • management
  • mathematics
  • sociology
  • state and public administration
by year
  • 2028
  • 2027
  • 2026
  • 2025
  • 2024
  • 2023
  • 2022
  • 2021
  • 2020
  • 2019
  • 2018
  • 2017
  • 2016
  • 2015
  • 2014
  • 2013
  • 2012
  • 2011
  • 2010
  • 2009
  • 2008
  • 2007
  • 2006
  • 2005
  • 2004
  • 2003
  • 2002
  • 2001
  • 2000
  • 1999
  • 1998
  • 1997
  • 1996
  • 1995
  • 1994
  • 1993
  • 1992
  • 1991
  • 1990
  • 1989
  • 1988
  • 1987
  • 1986
  • 1985
  • 1984
  • 1983
  • 1982
  • 1981
  • 1980
  • 1979
  • 1978
  • 1977
  • 1976
  • 1975
  • 1974
  • 1973
  • 1972
  • 1971
  • 1970
  • 1969
  • 1968
  • 1967
  • 1966
  • 1965
  • 1964
  • 1963
  • 1958
  • More
Subject
News
October 6, 2026
International N5 Symposium ‘Neural Networks and Nonlinearity in Nizhny Novgorod Brings Together Scientists from Russia and Serbia
The International N5 Symposium ‘Neural Networks and Nonlinearity in Nizhny Novgorod’ was held at the Nizhny Novgorod House of Scientists from September 23 to 26. The event was organised by HSE University–Nizhny Novgorod and the Nizhny Novgorod House of Scientists, with the participation of Sberbank and the Institute of Physics Belgrade. The symposium was held for the second time: the first conference took place in 2025 and attracted considerable interest from the academic community.
October 5, 2026
‘The Climate Transition Is Not Necessarily a Limitation for Business
Linara Khadimullina works in the field of low-carbon development. In an interview with the Young Scientists of HSE project, she spoke about why nature is not just a beautiful backdrop, her research on the role of sustainable corporate governance in reducing greenhouse gas emissions, and growing plants as a source of inspiration.
October 5, 2026
Africa, Youth, and Civic Dialogue: Public Diplomacy Discussed at HSE University
In late September, HSE University hosted a roundtable discussion titled Civil Society in African Countries and Youth Participation in Public Diplomacy. Representatives of non-governmental organisations from Ghana, Ethiopia, and Russia, along with students from HSE University’s Bachelor’s Programme in Public Administration, discussed how young people without official diplomatic status can influence relations between countries and how the nonprofit sector can remain sustainable amid declining grant funding.

 

Have you spotted a typo?
Highlight it, click Ctrl+Enter and send us a message. Thank you for your help!

Publications
  • Books
  • Articles
  • Chapters of books
  • Working papers
  • Report a publication
  • Research at HSE

?

The Use of Khislavichi Lect Morphological Tagging to Determine its Position in the East Slavic Group

P. 174–186.
Afanasev I.

The study of low-resourced East Slavic lects is becoming increasingly relevant as they face the prospect of extinction under the pressure of standard Russian while being treated by academia as an inferior part of this lect. The Khislavichi lect, spoken in a settlement on the border of Russia and Belarus, is a perfect example of such an attitude.We take an alternative approach and study East Slavic lects (such as Khislavichi) as separate systems. The proposed method includes the development of a tagged corpus through morphological tagging with the models trained on the bigger lects. Morphological tagging results may be used to place these lects among the bigger ones, such as standard Belarusian or standard Russian. The implemented morphological taggers of standard Russian and standard Belarusian demonstrate an accuracy higher than the accuracy of multilingual models by 3 to 15{%. The study suggests possible ways to adapt these taggers to the Khislavichi dataset, such as tagset unification and transcription closer to the actual sound rather than the standard lect pronunciation. Automatic classification supports the hypothesis that Khislavichi is a border East Slavic lect that historically was Belarusian but got russified: the algorithm places it either slightly closer to Russian or to Belarusian.

Language: English
Full text
Text on another site
Keywords: диалектологияавтоматическая обработка естественного языкаautomatic classificationавтоматическая классификацияdialectologymorphological taggingморфологическая разметкаdigital dialectologyцифровая диалектологияNatural Language Processing (NLP)KhislavichiХиславичи

In book

Proceedings of Tenth Workshop on NLP for Similar Languages, Varieties and Dialects (VarDial 2023)
Association for Computational Linguistics, 2023.
Similar publications
Lacuna Inc. at SemEval-2026 Task 4: Structurally Gated State-Space Models for Disentangling Narrative Similarity
Куделя А. В., Алшауи Р., Shirnin A., , in: Proceedings of the 20th International Workshop on Semantic Evaluation (2026).: San Diego: Association for Computational Linguistics, 2026. P. 2347–2353.
In this paper, we present the Invariant-Variant Disentangled State-Space Model (IVD-SSM),our submission to SemEval-2026 Task 4 on Narrative Story Similarity and Narrative Representation Learning. Evaluating narrative similarity is a profound computational challenge that requires models to look past concrete, superficial elements such as specific names, actors, objects, or settings to isolate and compareabstract patterns of ...
Added: September 28, 2026
The Classics at SemEval-2026 Task 3: Combining Transformer Models and LLM-Generated Annotations for Dimensional Aspect-Based Sentiment Analysis
Алшауи Р., Raj A., Куделя А. В. et al., , in: Proceedings of the 20th International Workshop on Semantic Evaluation (2026).: San Diego: Association for Computational Linguistics, 2026. P. 2648–2656.
This paper presents an approach to the SemEval-2026 Task 3: Dimensional Aspect-Based Sentiment Analysis. We investigate methods for moving beyond traditional categorical sentiment (e.g., positive or negative) to predict fine-grained, real-valued scores for sentiment “valence” (positivity) and “arousal” (intensity). We participate in two subtasks: predicting these scores for given aspects (Subtask 1) and extracting full ...
Added: September 28, 2026
Proceedings of the 20th International Workshop on Semantic Evaluation (2026)
San Diego: Association for Computational Linguistics, 2026.
Added: September 28, 2026
Акцентная система говора деревни Пятиусово: существительные а-склонения
Никитенко Д. Ю., RHEMA. РЕМА 2026 № 1 С. 188–207
This  study  explores  a  fragment  of  the  accentual  system  of  the  Pskov dialect,  with  data  obtained  from  both a  corpus  and  fieldwork.  The  focus is on a-stem nouns (1st declension), with the accent types and stress patterns  described. The accent types of nouns are presented in tables, while the stress patterns are given in a dictionary format, allowing for a combined approach to describing the dialect’s accentual system. Special attention is paid to stress placement  in  the  accusativus  singularis  form  of  a  number  of  words,  which in  this  case ...
Added: September 22, 2026
Lecture Notes in Artificial Intelligence
Springer, 2026.
Two volumes of the SPECOM 2026 proceedings contain a collection of submitted papers presented at SPECOM 2026, which were thoroughly reviewed by members of the Program Committee and additional reviewers consisting of almost 80 experts in the conference topic areas. In total, 65 regular full papers out of 99 submissions made via the EasyChair electronic ...
Added: September 20, 2026
Интернет-ресурсы по изучению миноритарных языков Ирана
Gromova A., В кн.: I Международная научно-образовательная конференция «Пейсиковские чтения: проблемы современного академического востоковедения»: материалы конференции.: М.: ИСАА МГУ имени М.В. Ломоносова, 2023. С. 15–20.
В проектах по ревитализации миноритарных языков одной из главных проблем остается подготовка педагогических кадров и учебных материалов. Размещение онлайн видеоуроков и диалектных аудиоархивов, наряду с другими формами проводимой «снизу» работы, направленной на корректировку текущей языковой ситуации, даёт угрожаемым идиомам новые возможности «быть увиденными и услышанными». ...
Added: June 30, 2026
Компьютерная лингвистика и интеллектуальные технологии: По материалам ежегодной международной конференции «Диалог». Выпуск 24
M.: Max press, 2026.
The volume includes 64 papers from the international conference on computational linguistics and intelligent technologies 'Dialogue 2026,' representing a broad spectrum of theoretical and applied research in the field of natural language description, language process modeling, and the development of practically applicable computational linguistic technologies. For specialists in theoretical and applied linguistics and intelligent technologies. ...
Added: June 27, 2026
Proceedings of the Sixth Workshop on Teaching NLP (TeachNLP 2024)
Association for Computational Linguistics, 2024.
Added: June 14, 2026
Analysis of Images, Social Networks and Texts. AIST 2024
Cham: Springer, 2024.
Added: June 14, 2026
Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 4: Student Research Workshop)
Association for Computational Linguistics, 2026.
Added: June 13, 2026
A textual fingerprint learning model to detect fake information spreaders in social networks
Behzadidoost R., Neurocomputing 2025 Vol. 665 P. 1–21
While earlier research has focused on detecting misinformation content, identifying the users who spread it, referred to in this paper as fake information spreaders, remains a relatively new challenge. These users deliberately mix true and false information, making detection more difficult. This paper proposes a textual fingerprint learning model to detect fake information spreaders. The ...
Added: March 12, 2026
Дискриминативная лемматизация сокращений в эпоху LLM
Глазкова А. В., Смаль И. В., Lyashevskaya O. et al., Доклады Российской академии наук. Математика, информатика, процессы управления (ранее - Доклады Академии Наук. Математика) 2025 Т. 527 С. 146–155
This paper presents a study on the effectiveness of discriminative methods for abbreviation lemmatization in Russian texts. Unlike generative approaches, discriminative models select the optimal lemma from a fixed set of candidates, eliminating the risk of generating grammatically incorrect word forms. For the first time in Russian language processing, we conduct a comprehensive analysis of ...
Added: March 10, 2026
Transformer-based approaches for lemmatizing abbreviations in Russian texts
Glazkova A., Lyashevskaya O., Morozov D. et al., Journal of Mathematical Sciences 2025 Vol. 546 P. 32–47
This paper addresses the task of lemmatizing abbreviations in the Russian language. Abbreviation lemmatization is particularly challenging, as it involves not only transforming a word into its normal form but also correctly expanding the abbreviation. We explore two approaches to this task, both leveraging large pretrained language models. The first approach is generative, where the ...
Added: March 10, 2026
Грамматический ландшафт художественной прозы: динамика частеречных распределений в русском рассказе XX века
Kirina M., В кн.: Русская грамматика: полипарадигмальность как методологический принцип современных научных исследований : материалы IX Международного научного симпозиума.: Издательство ИГУ, 2025. С. 270–275.
В статье представлены результаты пилотного исследования, направленного на описание дистрибуции частей речи в синхронии и диахронии на материале русской прозы малой формы. Рассматриваются изменения морфологического состава художественных текстов (на уровне грамматических классов) на протяжении XX века в соответствии с 9 историко-культурными периодами. Материалом исследования выступает выборка из 943 рассказов суммарным объемом более 3 млн. словоупотреблений. ...
Added: February 28, 2026
Образцы говора македонских переселенцев в Южном Банате Республики Сербии, сёла Качарево и Глогонь, община Панчево
Muravleva N., В кн.: Исследования по славянской диалектологии. Выпуск 25Т. 25.: М.: Институт славяноведения РАН, 2025. С. 426–441.
В статье публикуются нарративы на македонском языке, записанные во время экспедиции 2023 года (Борисов, Кикило, Немчинов 2024) у ин формантов — представителей македонского меньшинства, проживаю щих в сёлах Качарево и Глогонь (серб. Kačarevo, Glogonj) общины Пан чево, Воеводина, Республика Сербия. В диалектных текстах отражены контактные явления, возникшие под влиянием мажоритарного сербского языка, а также смешение ...
Added: February 18, 2026
Претериальные формы в идиоме македонских переселенцев Воеводины (Сербия)
Muravleva N., Славянский мир в третьем тысячелетии 2025 Т. 20 № 3-4 С. 144–172
The article examines the features of the past tense system in Macedonian resettlement dialects of the Autonomous Province of Vojvodina, Serbia, based on a corpus of texts collected during a 2023 linguistic expedition to the villages of Jabuka, Kačarevo, Glogonj, Plandište, and Belgrade. The first section provides a sociolinguistic overview of the formation of the ...
Added: February 18, 2026
Development of a Language Model for Automated Classification of English-Language Scientific Articles by SRSTI Codes
V. V. Zunin, A. I. Afonin, V. I. Anoshin et al., Automatic Documentation and Mathematical Linguistics 2025 Vol. 59 No. 5 P. 287–293
The development of an artificial intelligence-based language model for classifying English-language scientific articles by SRSTI codes is described. This improves the processes of reviewing and indexing scientific publications. A pre-processed dataset of scientific articles was used for training and testing the models. An architecture for cascade classification was developed, and the performance of models with ...
Added: February 11, 2026
30th International Conference on Applications of Natural Language to Information Systems, NLDB 2025, Kanazawa, Japan, July 4–6, 2025, Proceedings, Part I. Natural Language Processing and Information Systems. (LNCS, volume 15836)
Springer, 2025.
The two-volume set LNCS 15836 and 15837 constitutes the proceedings of the 30th International Conference on Applications of Natural Language to Information Systems, NLDB 2025, held in Kanazawa, Japan, during July 4–6, 2025. The 33 full papers, 19 short papers and 2 demo papers presented in this volume were carefully reviewed and selected from 120 submissions. ...
Added: February 3, 2026
  • About
  • About
  • Key Figures & Facts
  • Sustainability at HSE University
  • Faculties & Departments
  • International Partnerships
  • Faculty & Staff
  • HSE Buildings
  • HSE University for Persons with Disabilities
  • Public Enquiries
  • Studies
  • Admissions
  • Programme Catalogue
  • Undergraduate
  • Graduate
  • Exchange Programmes
  • Summer University
  • Summer Schools
  • Semester in Moscow
  • Business Internship
  • Research
  • International Laboratories
  • Research Centres
  • Research Projects
  • Monitoring Studies
  • Conferences & Seminars
  • Academic Jobs
  • Yasin (April) International Academic Conference on Economic and Social Development
  • Media & Resources
  • Publications by staff
  • HSE Journals
  • Publishing House
  • iq.hse.ru: commentary by HSE experts
  • Library
  • Economic & Social Data Archive
  • Video
  • HSE Repository of Socio-Economic Information
  • HSE1993–2026
  • Contacts
  • Copyright
  • Privacy Policy
  • Site Map
Edit