Development of an Improvised Technique for Classification in Data Mining

dc.contributor.guideIndu Kashyap
dc.coverage.spatial
dc.creator.researcherNair,Preeti
dc.date.accessioned2021-05-04T06:29:15Z
dc.date.available2021-05-04T06:29:15Z
dc.date.awarded2020
dc.date.completed2019
dc.date.registered2012
dc.description.abstractData mining is the process of finding useful hidden patterns from huge amount of data that are produced by information systems and using those patterns in smart decision making. The predefined methods and algorithms that are used to extract these useful patterns are called data mining techniques.Among these, classification is a supervised learning approach. Classification is a technique in which a given unknown instance is categorized into a particular class or label. An algorithm that implements classification, known as a classifier. A classifier analyses the training instances with labels and builds a model based on them to predict labels of test instances that are unknown. Classification algorithms are based on many learning methods such as instance-based, rule-based, tree-based etc.The k nearest neighbor (k NN) algorithm is one of the most widely used classification methods, which is a type of instance-based, non-parametric learning method or lazy learning method. The basis of k nearest neighbor classifiers is learning by finding similarities between instances. The similarities are measured by a distance formula.The motivation behind selecting k NN classification algorithm for this research work is that it has widely been used in applications of data mining and machine learning because of its inherent simplicity in implementation and note-worthy performance.In this research work, several conceptual variations of k NN have been proposed which lead to remarkable improvements in classification performance in various ways. In the first phase of this research, in order to enhance classification performance, an ensemble model has been constructed using stacking approach.Further enhancements in later phases led to certain limitations being eradicated. Rare class problems are one of the limitations of k NN as it is an instance-based non parametric learning method, which means there is no prior model assumption so it is much more sensitive to imbalanced data.
dc.description.noteClassification, k nearest neighbor, classification, imbalanced data, distance computation, accuracy, precision, recall, Fmeasure, ensemble method, IQR, SMOTE, resample method, hybrid, outliers, performance.
dc.format.accompanyingmaterialDVD
dc.format.dimensions
dc.format.extent
dc.identifier.urihttp://hdl.handle.net/10603/324265
dc.languageEnglish
dc.publisher.institutionDepartment of Computer Science Engineering
dc.publisher.placeFaridabad
dc.publisher.universityManav Rachna International University
dc.relation
dc.rightsuniversity
dc.source.universityUniversity
dc.subject.keywordComputer Science
dc.subject.keywordComputer Science Information Systems
dc.subject.keywordEngineering and Technology
dc.titleDevelopment of an Improvised Technique for Classification in Data Mining
dc.title.alternative
dc.type.degreePh.D.

Files

Original bundle

Now showing 1 - 5 of 18
Loading...
Thumbnail Image
Name:
01_title.pdf
Size:
110.92 KB
Format:
Adobe Portable Document Format
Description:
Attached File
Loading...
Thumbnail Image
Name:
02_declaration.pdf
Size:
67.77 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
03_certificate.pdf
Size:
79.15 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
04_acknowledgement.pdf
Size:
75.53 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
05_contents.pdf
Size:
33.66 KB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.79 KB
Format:
Plain Text
Description: