Using Generative Pretrained Transformer-3 Models for Russian News Clustering and Title Generation tasks

M. Tikhonova; Pisarevskaya D.; T. Shavrina; Shliazhko O.

doi:10.28995/2075-7182-2021-20-1214-1223

Publications

?

Using Generative Pretrained Transformer-3 Models for Russian News Clustering and Title Generation tasks

Komp'juternaja Lingvistika i Intellektual'nye Tehnologii. 2021. Vol. 20. P. 1214–1223.

Tikhonova M., Pisarevskaya D., Shavrina T., Shliazhko O.

The paper presents a methodology for news clustering and news headline generation based on the zero-shot
approach and minimal tuning of the RuGPT-3 architecture (Generative Pretrained Transformer 3 for Russian). The
solution is presented in a competition for news clustering, headline selection and generation.
The following approaches are described: 1) zero-shot unsupervised classification based on pairwise news
perplexity: the method requires no training or model fine-tuning and yields 0.7 F1-measure.
2) fine-tuning: news headlines generation with the best result 0.292 ROUGE and 0.596 BLEU.

Research target: Computer Science

Keywords: text generation Generative pretrained transformer Evaluation track text clustering ruGPT-3

Кодовые конструкции на базе обобщенных каскадных кодов для систем связи, использующих прием на основе порядковых статистик

Osipov D., Информационно-управляющие системы 2026 № 3 С. 49–62

Introduction: In many communication systems under construction and those to be created power control and channel estimation techniques developed for the previous generation communication systems fail to provide desired precision. One way to solve this problem is to use order-statistics-based reception techniques that do not need channel estimation or power control. To ensure the desired ...

Added: July 3, 2026

Влияние социально-демографических факторов на эмоциональное восприятие текста цифровой коммуникации: опыт экспериментального исследования

Герцен А. С., Виртуальная коммуникация и социальные сети 2026 Т. 5 С. 198–206

Emotional interpretation of digital texts requires the reconstruction of the author’s intent. According to Activity Theory, emotions are shaped by historical and cultural factors rather than human biology. Using the Geneva Emotion Wheel model, the authors studied responses to a digital communication text in order to measure the convergence between attributed emotions (those the reader ...

Added: July 2, 2026

The 12th International Conference on Information Technology and Quantitative Management (ITQM 2025)

Netherlands: ScienceDirect, 2025.

No ...

Added: June 28, 2026

Object-centric process management: A research manifesto

Seidel A., Weske M., Montali M. et al., Information Systems 2026 Vol. 141 Article 102728

Business process management employs process models and event logs to represent the behavior of the information systems under study. Traditional case-centric notions consider the order of activities and events in isolated process instances. The emerging field of object-centric processes challenges this assumption by putting objects in the center. Object-centric process mining and modeling approaches identify ...

Added: June 27, 2026

2024 26th International Conference on Digital Signal Processing and its Applications (DSPA)

IEEE, 2024.

A.S. Popov Russian Science and Technical Society with support from V. A. Trapeznikov Institute of Control Sciences, V.A. Kotelnikov Institute of Radio Engineering and Electronics, Autex Ltd. is leading the ХХVIII International Conference «Digital Signal Processing and its Applications — DSPA-2024» ...

Added: June 27, 2026

Построение методик оценки качества восприятия (QOE) потокового видео

Ivchenko A., Дворкович А. В., Телекоммуникации 2020 Т. 12 С. 2–11

Dynamic Adaptive Streaming over HTTP (DASH) technology powers most multimedia services. Its specific features (re-buffering, quality switching, etc.) necessitate the development of specialized methods for assessing user subjective quality of experience (QoE) based on objective parameters. This article examines the impact of various metrics on QoE and presents assessment models with Spearman correlation coefficients up ...

Added: June 27, 2026

Платформа, управляемая событиями, для интеграции компонентов машинного зрения с операционным центром.

Gadzhimirzaev S., Хельвас А. В., 2023 3rd International Conference on Innovative Research in Applied Science, Engineering and Technology (IRASET) Mohammedia, Morocco 2023 P. 1–6

The article proposes the architecture for eventdriven Emergency Operation Center with Machine Vision Component. Sources of information are analyzed and approaches to machine vision events for tactical situations detection and estimation are discussed. Messages from Machine Vision Components are converted to Common Alerting Protocol and processed by Operation Center environment for tactical situations recognition. ...

Added: June 26, 2026

Дискретное моделирование процесса восстановительного ремонта участка дороги

Gadzhimirzaev S., Хельвас А. В., Компьютерные исследования и моделирование 2022 Т. 14 № 6 С. 1255–1268

This work contains a description of the results of modeling the process of maintaining the readiness of a section of the road network under strikes of with specified parameters. A one-dimensional section of road up to 40 km long with a total number of strikes up to 100 during the work of the brigade is ...

Added: June 26, 2026

Подход к оценке динамики уровня консолидированности отрасли

Gadzhimirzaev S., Хельвас А. В., Лукьянченко П. П., Computer Research and Modeling 2023 Vol. 15 No. 1 P. 129–140

In this article we propose a new approach to the analysis of econometric industry parameters for the industry consolidation level. The research is based on the simple industry automatic control model. The state of the industry is measured by quarterly obtained econometric parameters from each industry’s company provided by the tax control regulator. An approach ...

Added: June 26, 2026

Цифровой двойник полностью автоматизированного склада с глубокими стеллажами

Gadzhimirzaev S., Хельвас А. В., International Frequency Sensor Association (IFSA) Publishing, 19-21 February 2025 Granada, Spain 2025 P. 172–176

The paper presents models for an innovative fully robotic warehouse for storing boxed goods. A discrete multiagent simulation of the movement of shuttles in a warehouse for a given sequence of pallet shipments has been implemented. Different strategies for placement of boxes in various areas of a warehouse are evaluated, as well as optimal routing ...

Added: June 26, 2026

Incorporating Scientific Knowledge into Neural Network Density Functionals

Medvedev M., Journal of Chemical Theory and Computation 2026 Vol. 22 No. 9

Density functional theory (DFT) is the workhorse of modern reactions and materials modeling. While the exact functional remains unknown, many approximations to it have been constructed either by hand-crafting functional forms to satisfy exact constraints or by machine learning. In this work, we show how both of these approaches can be fused to build both ...

Added: June 26, 2026

Моделирование полностью роботизированного склада со стеллажами глубокого хранения

Gadzhimirzaev S., Хельвас А. В., Computer Research and Modeling 2026 Vol. 18 No. 2 P. 423–438

This article presents a model of a fully automated warehouse with deep storage racks designed for boxed goods storage. The study focuses on optimizing warehouse operations through discrete multiagent simulation of shuttle movements for pallet loading and unloading tasks. The authors investigate various product placement strategies, including the Nearest Channel Positioning Algorithm (NCPA), Most Empty ChannelGroup Placement (MECGP), and ...

Added: June 24, 2026

A machine learning dataset on winter roads of Krasnoyarsk Krai, Russia for the forestry and infrastructural projects

Podolskaia E., Sinitsina A., European Journal of Forest Engineering 2026 Vol. 12 No. 1 P. 7–21

Machine learning in transport modeling has become a trend in science and industry. In this paper, we observe its main directions and focus on a dataset of seasonal road creation. Seasonality as a parameter in transport modeling has a significant impact on transport scenarios but is underestimated worldwide and in Russia, despite modern data challenges. ...

Added: June 24, 2026

The state and prospects of using virtual reality technologies in sports: a brief review

Atlasov B., Selskiy A., Russian Journal of Information Technology in Sports 2025 Vol. 2 No. 1 P. 13–21

The article examines the current state of the global virtual and augmented reality (VR/AR) technology market in sports, noting its growth, although slower than previously expected. Special attention is paid to the Russian market, where the development of VR technologies in sports lags behind world leaders such as the United States, EU countries and China, ...

Added: June 23, 2026

AI & PDE: ICLR 2026 Workshop on AI and Partial Differential Equations

[б.и.], 2026.

Added: June 23, 2026

Alibaba и Open Source. История и масштабы сотрудничества китайской корпорации и мира открытого кода.

Silakov D., Системный администратор 2026 № 4 С. 38–43

Alibaba Group – китайский гигант электронной коммерции – владелец маркетплейсов AliExpress, Taobao и Tmall, платежной системы AliPay, а также крупнейшего в КНР сервиса облачных вычислений – Alibaba Cloud. В последние годы внимание к компании приковано благодаря ее достижениям в области искусственного интеллекта – технологии Tongyi Qianwen и открытых моделей линейки Qwen, доступной всем желающим. Но ...

Added: June 23, 2026

2025 9th International Conference on Information, Control, and Communication Technologies (ICCT-2025)

IEEE, 2026.

The 9th International Scientific Conference on Information, Control, and Communication Technologies (ICCT-2025) had been held October 7-11, 2025 in Gomel, Belarus. The main technical areas and applications covered by the proceedings are optoelectronics, acousto-optic, microwave technology, antenna systems, measuring technology, metamaterials, nanostructures, nanofilms, photonic crystals, biology and medicine, biophotonics, bioengineering, neural networks in communication technologies; ...

Added: June 23, 2026

Proceedings of the 4th Workshop on NLP for Music and Audio (NLP4MusA 2026)

Buzaev F., Mullakhmetov R., Bogachev R. et al., Association for Computational Linguistics, 2026.

Playlist generation based on textual queries using large language models (LLMs) is becoming an important interaction paradigm for music streaming platforms. User queries span a wide spectrum from highly personalized intent to essentially catalog-style requests. Existing systems typically rely on non-personalized retrieval/ranking or apply a fixed level of preference conditioning to every query, which can ...

Added: June 22, 2026

Zα and Zβ Localize ADAR1 to Flipons That Modulate Innate Immunity, Alternative Splicing, and Nonsynonymous RNA Editing

Herbert A., Cherednichenko O., Lybrand T. et al., International Journal of Molecular Sciences 2025 Vol. 26 No. 6 Article 2422

The double-stranded RNA editing enzyme ADAR1 connects two forms of genetic programming, one based on codons and the other on flipons. ADAR1 recodes codons in pre-mRNA by deaminating adenosine to form inosine, which is translated as guanosine. ADAR1 also plays essential roles in the immune defense against viruses and cancers by recognizing left-handed Z-DNA and ...

Added: June 22, 2026

Международная конференция «Математические идеи академика П.Л. Чебышёва, их приложения в естественных науках и технологи- ях искусственного интеллекта», приуроченная к 205-й годовщине со дня его рождения» : Материалы конференции. / (Обнинск, 14–16 мая 2026 г.): Материалы конференции. Под ред. акад. В.Б. Бетелина. — Калуга: Калужский печатный двор, 2026. — 232 с.

Калужский печатный двор, 2026.

Conference Proceedings INTERNATIONAL CONFERENCE “Mathematical Ideas of Academician P.L. Chebyshev, Their Applications in Natural Sciences and Artificial Intelligence Technologies” dedicated to the 205th anniversary of his birth ...

Added: June 20, 2026

ИНТЕГРАЦИЯ ТЕХНОЛОГИИ ГЕНЕРАТИВНОГО ИСКУССТВЕННОГО ИНТЕЛЛЕКТА В ОБРАЗОВАТЕЛЬНЫЙ ВИДЕОКОНТЕНТ

Stognieva O., Чеснокова Н. Е., Отечественная и зарубежная педагогика 2026 Т. 1 № 3 (115) С. 123–131

Integration of generative artificial intelligence tools into educational practice highlights the need for pedagogically grounded approaches to their use in the creation of educational video content, which is increasingly applied in language and professionally oriented instruction. The purpose of this article is to conduct a comparative analysis of educational video content created using generative AI tools ...

Added: June 20, 2026

Beyond Delta: Introducing an Angle Metric for Stylometric Similarity

Gorina O. G., / Series rs-8457236 "Research Square Preprints". 2025.

This paper addresses the problem of measuring stylometric similarity between texts. Traditional methods, such as Burrows’s Delta and its modifications, have several limitations, including dependence on a reference corpus and sensitivity to sample size, which reduces their reliability when working with short texts of fewer than 5,000 words. As an alternative, we propose a novel ...

Added: June 8, 2026

TEncDM: Understanding the Properties of the Diffusion Model in the Space of Language Model Encodings

Shabalin A., Meshchaninov V., Chimbulatov E. et al., , in: Proceedings of the 39th Annual AAAI Conference on Artificial IntelligenceVol. 39. Issue 23.: Washington, United States of America: AAAI Press, 2025. Ch. 110 P. 25110–25118.

This paper presents the Text Encoding Diffusion Model (TEncDM), a novel approach to diffusion modeling that operates in the space of pre-trained language model encodings. In contrast to traditionally used embeddings, encodings integrate contextual information. In our approach, we also employ a transformer-based decoder, specifically designed to incorporate context in the token prediction process. We ...

Added: December 18, 2025

Искусственный интеллект как симулякр смысла

Малинов С. А., Галактика медиа: журнал медиа исследований 2025 Т. 7 № 4 С. 154–173

In recent years, artificial intelligence (AI) has been actively integrated into everyday human life. Its popularity continues to grow steadily, and companies increasingly employ AI to optimize and accelerate workflows. Ordinary users leverage large language models (LLMs) and multimodal AI systems to perform a wide range of tasks, including generating texts, images, and videos; planning ...

Added: December 7, 2025