Scinovex
articleTop 1% cited

Data mining: concepts and techniques

Choice Reviews Online · 2012 · Vol. 49(06) · pp. 49–3305
Jiawei HanMicheline Kamber

Abstract

Our ability to generate and collect data has been increasing rapidly. Not only are all of our business, scientific, and government transactions now computerized, but the widespread use of digital cameras, publication tools, and bar codes also generate data. On the collection side, scanned text and image platforms, satellite remote sensing systems, and the World Wide Web have flooded us with a tremendous amount of data. This explosive growth has generated an even more urgent need for new techniques and automated tools that can help us transform this data into useful information and knowledge.Like the first edition, voted the most popular data mining book by KD Nuggets readers, this book explores concepts and techniques for the discovery of patterns hidden in large data sets, focusing on issues relating to their feasibility, usefulness, effectiveness, and scalability. However, since the publication of the first edition, great progress has been made in the development of new data mining methods, systems, and applications. This new edition substantially enhances the first edition, and new chapters have been added to address recent developments on mining complex types of data- including stream data, sequence data, graph structured data, social network data, and multi-relational data.Whether you are a seasoned professional or a new student of data mining, this book has much to offer you:* A comprehensive, practical look at the concepts and techniques you need to know to get the most out of real business data.* Updates that incorporate input from readers, changes in the field, and more material on statistics and machine learning.* Dozens of algorithms and implementation examples, all in easily understood pseudo-code and suitable for use in real-world, large-scale data mining projects.* Complete classroom support for instructors at www.mkp.com/datamining2e companion site.

Citations
28,852
FWCI
1742.71
field-weighted impact
References
487
Percentile
100%
vs. same field & year
Citations per year
Cited by
Enhancing cyber security by predicting malwares using supervised machine learning models
International Journal of Computing and Artificial Intelligence · 2021 · 10 citations
An experimental approach for discovering association rules using FP-growth algorithm
International Journal of Circuit Computing and Networking · 2021 · 0 citations
An ensemble classification approach for prediction of banknote authentication
International Journal of Computing Programming and Database Management · 2021 · 0 citations
An experimental approach for discovering frequent patterns
International Journal of Circuit Computing and Networking · 2021 · 0 citations
Diagnosis spinal abnormalities utilizing machine learning algorithms
International Journal of Computing Programming and Database Management · 2021 · 0 citations
A decision tree method for building energy demand modeling
Energy and Buildings · 2010 · 605 citations
Black hole: A new heuristic optimization approach for data clustering
Information Sciences · 2012 · 1,346 citations
References
Neural networks for pattern recognition
Choice Reviews Online · 1994 · 18,690 citations
Classification and Regression Trees.
Biometrics · 1984 · 23,850 citations
SPADE: An Efficient Algorithm for Mining Frequent Sequences
Machine Learning · 2001 · 1,811 citations
Greedy function approximation: A gradient boosting machine.
The Annals of Statistics · 2001 · 27,794 citations
Algorithms for Clustering Data
Technometrics · 1990 · 7,836 citations
Procedures for Detecting Outlying Observations in Samples
Technometrics · 1969 · 3,689 citations
Measuring the Accuracy of Diagnostic Systems
Science · 1988 · 9,901 citations
Data clustering
ACM Computing Surveys · 1999 · 13,065 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.