International Journal of Science and Research (IJSR)

International Journal of Science and Research (IJSR)
Call for Papers | Fully Refereed | Open Access | Double Blind Peer Reviewed

ISSN: 2319-7064


Downloads: 122

Research Paper | Computer Science & Engineering | India | Volume 2 Issue 11, November 2013


Exploration of Data Mining Techniques in Record Deduplication

R. Gayathri [9] | A. Malathi [3]


Abstract: In todays business world, the database plays a vital role in decision making. As the organization grows, the size of the database also gets increased. This enormous growth in the database size leads to a problem of dirty data. Dirty data is the replicated data in the database which causes some issues like performance degradation, increasing operational cost and the lack of quality. This can be removed by the process of record deduplication. The record deduplication refers to identifying the same entity with different representations. Further cleaning and removing of replica in the repository become a mandatory work. Thus this paper surveys some of the record deduplication approaches. Also it compares with three approaches to record deduplication such as genetic programming, Modified BAT algorithm, and firefly algorithm approach with its limitation and advantages on all the three got discussed.


Keywords: Record Deduplication, preprocessing, Cleaning, Dirty data, genetic programming, mbat algorithm, firefly algorithm


Edition: Volume 2 Issue 11, November 2013,


Pages: 216 - 219


How to Download this Article?

You Need to Register Your Email Address Before You Can Download the Article PDF


How to Cite this Article?

R. Gayathri, A. Malathi, "Exploration of Data Mining Techniques in Record Deduplication", International Journal of Science and Research (IJSR), Volume 2 Issue 11, November 2013, pp. 216-219, https://www.ijsr.net/get_abstract.php?paper_id=02013451

Similar Articles with Keyword 'preprocessing'

Downloads: 0

Research Paper, Computer Science & Engineering, India, Volume 12 Issue 8, August 2023

Pages: 509 - 514

Enhancing Lung Cancer Detection with Deep Learning: A CT Image Classification Approach

Jeevika K S | Dr. Savitha S K

Share this Article

Downloads: 2 | Weekly Hits: ⮙1 | Monthly Hits: ⮙1

Research Paper, Computer Science & Engineering, India, Volume 11 Issue 10, October 2022

Pages: 660 - 667

Prediction of Student Admission using Fuzzy based Education Data Mining

Dr. Nikhat Khan

Share this Article
Top