Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Multivariate similarity-based conformity measure (MSCM: An outlier detection measure for data mining applications
University of Western Ontario.
German University in Cairo (GUC).ORCID iD: 0000-0003-4250-4752
Faculty of Science, Alexandria University.
2008 (English)In: Proceedings of the IASTED International Conference on Artificial Intelligence and Applications Machine Learning: as part of the 26th IASTED International Multi-Conference on Applied Informatics ; February 11 - 13, 2008, Innsbruck, Austria / [ed] Alexander Gammermann, Anaheim, Calif: ACTA Press, 2008, p. 314-320Conference paper, Published paper (Refereed)
Abstract [en]

Outliers, the odd objects in the dataset, can be viewed from two different perspectives; the outliers as undesirable objects that should be treated or deleted in the data preparation step of the data mining process, and the outliers as interesting objects that are identified for their own interest in the data mining step of the mining process. In the latter case, outliers shouldn't be removed, that's why one of the main categories of tasks performed by data mining techniques is outlier detection. Applications that make use of such detection include credit card fraud detection and network intrusion detection. Most of the available outlier detection techniques rely in a distance measure to compare the objects in the dataset which imposed the restriction of dealing with numeric data. In this paper a new multivariate similarity-based conformity measure (MSCM) is suggested to be used to detect outliers in datasets that contain attributes of different data types. The MSCM satisfies two other desirable features; being a multivariate measure and giving ranking instead of a binary judgment of the object. The measure has been applied on three different datasets in order to be evaluated; the measure has shown good results in these experiments.

Place, publisher, year, edition, pages
Anaheim, Calif: ACTA Press, 2008. p. 314-320
National Category
Information Systems, Social aspects
Research subject
Information systems
Identifiers
URN: urn:nbn:se:ltu:diva-27045Scopus ID: 2-s2.0-62849114512Local ID: 05a9f175-964a-4674-96fd-45e52b0fdab7ISBN: 9780889867093 (print)OAI: oai:DiVA.org:ltu-27045DiVA, id: diva2:1000226
Conference
IASTED International Conference on Artificial Intelligence and Applications Machine Learning : 11/02/2008 - 13/02/2008
Note

Upprättat; 2008; 20150625 (andbra)

Available from: 2016-09-30 Created: 2016-09-30 Last updated: 2021-03-11Bibliographically approved

Open Access in DiVA

No full text in DiVA

Scopus

Authority records

Elragal, Ahmed

Search in DiVA

By author/editor
Elragal, Ahmed
Information Systems, Social aspects

Search outside of DiVA

GoogleGoogle Scholar

isbn
urn-nbn

Altmetric score

isbn
urn-nbn
Total: 202 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf