
Recent Posts
 Write more papers or write better papers? (quantity vs quality)
 Using LaTeX for writing research papers
 An introduction to frequent subgraph mining
 We are launching a new data mining journal
 What is the job of a university professor?
 Introduction to clustering: the KMeans algorithm (with Java code)
 Happy New Year!
 Introduction to time series mining with SPMF
 Plagiarism at Ilahia College of Engineering and Technology by Nasreen Ali A and Arunkumar M
 Postdoctoral position in data mining in Shenzhen, China (apply now)
Categories
 Academia (5)
 artificial intelligence (4)
 Big data (21)
 Conference (12)
 Data Mining (51)
 Data science (17)
 General (25)
 Mathematics (2)
 Opensource (6)
 Plagiarism (1)
 Programming (15)
 Research (47)
 Time series (1)
 Uncategorized (1)
 Utility Mining (2)
Tag cloud
academic journal algorithm algorithms apriori articles big data characteristics classification comparison conference data mining data science datasets evaluation fpgrowth frequent pattern mining graph internet itemset mining java journal library M.Sc. map minsup opensource paper papers pattern mining peerreview Ph.D. prefixspan programmer programming publications Research research advisor research papers sequential patterns software source code spmf thesis topic website writingArchives
Recent Comments
Number of visitors:
405461
Tag Archives: data mining
Big Problems only found in Big Data?
Today, I will discuss the topic of Big Data, which is a very popular topic nowadays. The popularity of big data can be seen for example in universities. Many universities are currently searching for professors who do research on “big data”. Moreover, … Continue reading
Discovering and visualizing sequential patterns in web log data using SPMF and GraphViz
Today, I will show how to use the opensource SPMF data mining software to discover sequential patterns in web log data. Then, I will show to how visualize the frequent sequential patterns found, using GraphViz. Step 1 : getting the … Continue reading
Posted in Big data, Data Mining, Data science, Programming
Tagged data mining, graph, patterns, sequential patterns, spmf, visualization
8 Comments
Brief report about the ADMA 2013 conference
In this blog post, I will discuss my recent trip to the ADMA 2013 conference (9th Intern. Conf. on Advanced Data Mining and Applications in China (December 1416 2013 in Hangzhou, China at Zhejiang University). Note that the view expressed … Continue reading
How to encourage data mining researchers to share their source code and datasets?
A few months ago, I wrote a popular blog post on this blog about why it is important to publish source code and datasets for researchers“. I explained several advantages that researchers can get by sharing the source code of … Continue reading
Posted in Data Mining, Research
Tagged data mining, dataset, opensource, source code
Leave a comment
The importance of constraints in data mining
Today, I will discuss an important concept in data mining which is the use of constraints. Data mining is a broad field incorporating many different kind of techniques for discovering unexpected and new knowledge from data. Some main data mining … Continue reading
How to measure the memory usage of data mining algorithms in Java?
Today, I will discuss the topic of accurately evaluating the memory usage of data mining algorithms in Java. I will share several problems that I have discovered with memory measurements in Java for data miners and strategies to avoid these … Continue reading
Posted in Data Mining, Programming, Research
Tagged comparison, data mining, experiment, java, memory, performance
1 Comment
What are the steps to implement a data mining algorithm?
In this post, I will discuss what are the steps that I follow to implement a data mining algorithm. The subject of this post comes from a question that I have received by email recently, and I think that it … Continue reading
Posted in Data Mining, Programming
Tagged algorithm, data mining, design, implementation, programming
46 Comments
Analyzing the source code of the SPMF data mining software
Hi everyone, In this blog post, I will discuss how I have applied an opensource tool that is named Code Analyzer ( http://sourceforge.net/projects/codeanalyzegpl/ ) to analyze the source code of my opensource data mining software named SPMF. I have applied … Continue reading
How to characterize and compare data mining algorithms?
Hi, today, I will discuss how to compare data mining algorithms. This is an important question for data mining researchers who want to evaluate which algorithm is “better” in general or for a given situation. This question is also important … Continue reading
Posted in Data Mining, Programming, Research
Tagged algorithms, characteristics, classification, comparison, data mining, evaluation
7 Comments
A Map of Data Mining Algorithms (offered in SPMF v092c)
Hi, I have made a map to visualize the relationship between the 52 different data mining algorithms offered in the SPMF data mining software. You can view it in PNG format by clicking on the picture below: Or you can … Continue reading
Posted in Data Mining, Programming
Tagged algorithms, data mining, java, map, opensource, spmf
2 Comments