Abstrakt
Discretisation is a processing step often included in the preliminary data preparation. Typically, when the input features have continuous domains and their discrete forms are needed, all are translated into categorical type at the same time, before data mining takes place. However, proceeding this way is not always the most advantageous to performance. The paper presents results from the research where the discretisation transformations were carried out sequentially forward for variables, and their selection was based on their values and also importance of the attributes estimated by the constructed rankings. The experiments were executed on the datasets from the area of stylometric analysis of texts, the application domain focused on recognising authorship based on individual characteristics of writing styles. For the selected data mining techniques, the performance was studied in the context of transformed features. The observed trends indicate that along with enhanced understanding of the nature of the data, partial discretisation of feature sets could bring higher accuracy than transformation of entire input domain, showing the merits of the described research methodology.
| Język oryginału | angielski |
|---|---|
| Numer artykułu | 2679 |
| Czasopismo | Applied Sciences (Switzerland) |
| Tom | 16 |
| Numer wydania | 6 |
| Identyfikatory DOI | |
| Status publikacji | Opublikowano - mar 2026 |
Obszary tematyczne ASJC Scopus
- Materiałoznawstwo ogólne
- Instrumentacja
- Inżynieria ogólna
- Chemia i technologia procesów
- Zastosowania informatyki
- Procesy przepływu i przenoszenia płynów
Fingerprint
Zanurz się w tematy badawcze publikacji „Does All or Nothing Always Work Best? In Search of Advantageous Representation of Attributes”. Razem tworzą niepowtarzalny odcisk palca.Cytowanie
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver