Abstract
Concurrency is the ability of a system to allow for several operations to be carried out simultaneously. The ability to offer concurrency is unique to databases. This is one of the main properties that separates a database from other forms of data storage like spreadsheets.Duplicate detection lets organizations set duplicate detection policies and create duplicate detection rules for business and custom entities. These rules can be applied across different record types which allows for detection of fraud that takes place while rating a product in websites. For example, an organization may define that a lead is a duplicate of a contact, if they have the same name, same IP address, same ID. Based on the duplicate detection rules set by the administrator, the system alerts the admin about potential duplicates. To maintain data quality, you can schedule a duplicate detection job to check for duplicates for all records that match a certain criteria. You can clean the data by deleting, deactivating or merging the duplicates reported by a duplicate detection job. The proposed system adds more strength for detection of duplication by using ranking methodology.
Keywords
Duplicate detection
Data cleaning
Record linkage
Data deduplication.
Authors
How to Cite this Article
B.Praveen Kumar, M.R.Priyanka, C.Radhika, A.Subashini (2016).
"CONCURRENT PROGRESSIVE APPROACH FOR DETECTING THE DUPLICATION".
International Journal of Contemporary Research in Computer Science and Technology,
2(3), pp. 500-504.