Unknown

Dataset Information

0

Relevance, redundancy, and complementarity trade-off (RRCT): A principled, generic, robust feature-selection tool.


ABSTRACT: We present a new heuristic feature-selection (FS) algorithm that integrates in a principled algorithmic framework the three key FS components: relevance, redundancy, and complementarity. Thus, we call it relevance, redundancy, and complementarity trade-off (RRCT). The association strength between each feature and the response and between feature pairs is quantified via an information theoretic transformation of rank correlation coefficients, and the feature complementarity is quantified using partial correlation coefficients. We empirically benchmark the performance of RRCT against 19 FS algorithms across four synthetic and eight real-world datasets in indicative challenging settings evaluating the following: (1) matching the true feature set and (2) out-of-sample performance in binary and multi-class classification problems when presenting selected features into a random forest. RRCT is very competitive in both tasks, and we tentatively make suggestions on the generalizability and application of the best-performing FS algorithms across settings where they may operate effectively.

SUBMITTER: Tsanas A 

PROVIDER: S-EPMC9122960 | biostudies-literature | 2022 May

REPOSITORIES: biostudies-literature

altmetric image

Publications

Relevance, redundancy, and complementarity trade-off (RRCT): A principled, generic, robust feature-selection tool.

Tsanas Athanasios A  

Patterns (New York, N.Y.) 20220331 5


We present a new heuristic feature-selection (FS) algorithm that integrates in a principled algorithmic framework the three key FS components: relevance, redundancy, and complementarity. Thus, we call it relevance, redundancy, and complementarity trade-off (RRCT). The association strength between each feature and the response and between feature pairs is quantified via an information theoretic transformation of rank correlation coefficients, and the feature complementarity is quantified using pa  ...[more]

Similar Datasets

| S-EPMC5209828 | biostudies-literature
| S-EPMC4029432 | biostudies-literature
| S-EPMC1569877 | biostudies-literature
| S-EPMC11410923 | biostudies-literature
| S-EPMC6157248 | biostudies-literature
| S-EPMC7397300 | biostudies-literature
| S-EPMC5331171 | biostudies-literature
| S-EPMC7469506 | biostudies-literature
| S-EPMC10114348 | biostudies-literature
2015-11-30 | GSE72909 | GEO