Results 31 to 40 of about 10,119 (225)
Full Support for Efficiently Mining Multi-Perspective Declarative Constraints from Process Logs
Declarative process management has emerged as an alternative solution for describing flexible workflows. In turn, the modelling opportunities with languages such as Declare are less intuitive and hard to implement.
Christian Sturm +2 more
doaj +1 more source
D 3 -MapReduce: Towards MapReduce for Distributed and Dynamic Data Sets [PDF]
International audienceSince its introduction in 2004 by Google, MapRe-duce has become the programming model of choice for processing large data sets.
Anjos, Julio +12 more
core +4 more sources
Big data cleaning modeling of operation status of coal mine fully—mechanized coal mining equipment
In view of problems of large amount of data and noise and missing values existed in data of operation status of coal mine fully—mechanized coal mining equipment, a big data cleaning model of operation status of coal mine fully—mechanized coal mining ...
MA Hongwei +4 more
doaj +1 more source
Scalable Computation of Topological Abstractions for Scalar Data
Abstract Topological data analysis has become an important tool for large scale scalar data analysis and visualization, efficiently extracting the inherent structure and features of interest of the data. However, with growing dataset sizes and complexity, it is increasingly becoming infeasible to compute topological abstractions of interest in serial ...
M. Will +6 more
wiley +1 more source
Accumulative Computation on MapReduce [PDF]
MapReduce programming model attracts a lot of enthusiasm among both industry and academia, largely because it simplifies the implementations of many data parallel applications.
Liu, Yu +3 more
openaire +4 more sources
Evaluating MapReduce for seismic data processing using a practical application
Huge amounts of seismic data undergo complex iterative processing in the oil industry to get knowledge of the earth’s subsurface structure to detect where oil can be found and recovered.To evaluate the suitability of MapReduce for seismic processing ...
Chang-hai ZHAO +4 more
doaj +2 more sources
Parallel Computation of Rough Set Approximations in Information Systems with Missing Decision Data
The paper discusses the use of parallel computation to obtain rough set approximations from large-scale information systems where missing data exist in both condition and decision attributes.
Thinh Cao +4 more
doaj +1 more source
A Bibliometric Analysis of Process Mining
Process mining studies with numbers. ABSTRACT Process mining (PM) has emerged as a pivotal discipline in data science, bridging traditional process analysis with data‐driven techniques to extract actionable insights from event logs. This study conducts a comprehensive bibliometric analysis of 1764 peer‐reviewed articles from the Web of Science database
Seyfullah Tokumaci +2 more
wiley +1 more source
Spatial hotspot detection using polygon propagation
Spatial scan statistics is one of the most important models in order to detect high activity or hotspots in real world applications such as epidemiology, public health, astronomy and criminology applications on geographic data. Traditional scan statistic
Satya Katragadda +2 more
doaj +1 more source
PUC: parallel mining of high-utility itemsets with load balancing on spark
Distributed programming paradigms such as MapReduce and Spark have alleviated sequential bottleneck while mining of massive transaction databases. Of significant importance is mining High Utility Itemset (HUI) that incorporates the revenue of the items ...
Brahmavar Anup Bhat +2 more
doaj +1 more source

