?
Ускорение объединения распределенных наборов данных по заданному критерию
.
Tyryshkina Y.
Language:
Russian
In book
Мн.: [б.и.], 2022.
Tyryshkina Y., Tumkovskiy S., Информационно-управляющие системы 2022 № 5(120) С. 2–11
Added: May 31, 2022
Tyryshkina Y., , in: Proceedings of 2022 IEEE Moscow Workshop on Electronic and Networking Technologies (MWENT).: M.: IEEE, 2022.
Added: May 31, 2022
Tyryshkina Y., В кн.: Межвузовская научно-техническая конференция студентов, аспирантов и молодых специалистов имени Е.В. Арменского. Материалы конференции.: М.: МИЭМ НИУ ВШЭ, 2022.
В данной работе рассматривается проблема снижения затрат машинного времени за счет разработки и реализации метода ускорения операции соединения распределенных массивов данных по заданному критерию. Были решены следующие задачи: проведено исследование архитектуры распределенных хранилищ данных и алгоритмов параллельных вычислений; на основании этих исследований установлены лимитирующие стадии, замедляющие процесс переработки; разработан метод, исключающий установленные лимитирующие стадии; на ...
Added: May 31, 2022
С.Д. Кузнецов, Посконин А. В., Труды Института системного программирования РАН 2013 Т. 24 С. 327–258
Many modern applications (such as large-scale Web-sites, social networks, research projects, business analytics, etc.) have to deal with very large data volumes (also referred to as “big data”) and high read/write loads. These applications require underlying data management systems to scale well in order to accommodate data growth and increasing workloads. High throughput, low latencies ...
Added: January 30, 2018
Клеменков П. А., Kuznetsov S. D., Труды Института системного программирования РАН 2012 Т. 23 С. 143–158
Big data challenged traditional storage and analysis systems in several new ways. In this paper we try to figure out how to overcome this challenges, why it's not possible to make it efficiently and describe three modern approaches to big data handling: NoSQL, MapReduce and real-time stream processing. The first section of the paper is ...
Added: October 31, 2017
Shugurov I., Mitsyuk A. A., Proceedings of the Institute for System Programming of the RAS 2016 Vol. 28 No. 3 P. 103–122
Process mining is a relatively new research field, offering methods of business processes analysis and improvement, which are based on studying their execution history (event logs). Conformance checking is one of the main sub-fields of process mining. Conformance checking algorithms are aimed to assess how well a given process model, typically represented by a Petri ...
Added: September 12, 2016
Shmid A., Pozin B., Агейкин М. А. et al., М.: Пальмир, 2016.
The book introduces the latest technologies of processing big data (Big Data) on the example of the IBM BIG DATA platform ( underlying technology of an expert system IBM Watson). ...
Added: May 27, 2016
Зудин С., Gnatyshak D. V., Ignatov D. I., , in: Proceedings of the Twelfth International Conference on Concept Lattices and Their Applications Clermont-Ferrand, France, October 13-16, 2015Vol. 1466.: Clermont-Ferrand: CEUR Workshop Proceedings, 2015. P. 47–58.
In our previous work an efficient one-pass online algorithm
for triclustering of binary data (triadic formal contexts) was proposed.
This algorithm is a modified version of the basic algorithm for OAC-triclustering
approach; it has linear time and memory complexities. In
this paper we parallelise it via map-reduce framework in order to make
it suitable for big datasets. The results of ...
Added: October 23, 2015
Леохин Ю.Л., Мягков А.С., Информатизация образования и науки 2014 Т. 24 № 4 С. 111–118
In the article the implementation of the parallel programming model MapReduce, which is used in distributed computation, is considered. The results of the scalability research of the implementation running on the Plan9 operation system are shown. In addition, we have pointed out the main lines of the selected prototype development. ...
Added: October 23, 2014