Abstrakt
Discretisation often constitutes a part of initial data preparation stage. It translates continuous domain of features into granular, by assigning a number of intervals to represent attributes' values by nominal categories. Typically all real-valued features are subjected to transformations, regardless of their characteristics. The paper presents research on discretisation executed with a discerning approach. To all available attributes, feature selection mechanisms were employed, in the form of rankings that order variables based on their importance. Exploiting this discovered knowledge on attributes, discretisation was then driven by a ranking, and either highest or lowest ranking features were selected for transformation. The influence of selective discretisation on the performance of classification systems was studied for three popular inducers. The procedure was employed in the field of stylometry, and a task of authorship recognition, considered as a binary classification with balanced classes. The experiments show that discretisation based on importance of features can lead to better performance than in the case of transformations applied to all attributes.
| Język oryginału | angielski |
|---|---|
| Strony (od–do) | 3335-3344 |
| Liczba stron | 10 |
| Czasopismo | Procedia Computer Science |
| Tom | 176 |
| Identyfikatory DOI | |
| Status publikacji | Opublikowano - 2020 |
| Wydarzenie | 24th KES International Conference on Knowledge-Based and Intelligent Information and Engineering Systems, KES 2020 - Virtual Online Czas trwania: 16 wrz 2020 → 18 wrz 2020 |
Obszary tematyczne ASJC Scopus
- Informatyka ogólna
Fingerprint
Zanurz się w tematy badawcze publikacji „Performance evaluation for ranking-based discretisation”. Razem tworzą niepowtarzalny odcisk palca.Cytowanie
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver