Abstract:When machine learning supports decision-making in safety-critical systems, it is important to verify and understand the reasons why a particular output is produced. Although feature importance calculation approaches assist in interpretation, there is a lack of consensus regarding how features' importance is quantified, which makes the explanations offered for the outcomes mostly unreliable. A possible solution to address the lack of agreement is to combine the results from multiple feature importance quantifiers to reduce the variance of estimates. Our hypothesis is that this will lead to more robust and trustworthy interpretations of the contribution of each feature to machine learning predictions. To assist test this hypothesis, we propose an extensible Framework divided in four main parts: (i) traditional data pre-processing and preparation for predictive machine learning models; (ii) predictive machine learning; (iii) feature importance quantification and (iv) feature importance decision fusion using an ensemble strategy. We also introduce a novel fusion metric and compare it to the state-of-the-art. Our approach is tested on synthetic data, where the ground truth is known. We compare different fusion approaches and their results for both training and test sets. We also investigate how different characteristics within the datasets affect the feature importance ensembles studied. Results show that our feature importance ensemble Framework overall produces 15% less feature importance error compared to existing methods. Additionally, results reveal that different levels of noise in the datasets do not affect the feature importance ensembles' ability to accurately quantify feature importance, whereas the feature importance quantification error increases with the number of features and number of orthogonal informative features.

FeatureLTE: Learning to Estimate Feature Importance

The Berkelmans-Pries Feature Importance Method: A Generic Measure of Informativeness of Features

AutoFIS

FiBiNET: Combining Feature Importance and Bilinear feature Interaction for Click-Through Rate Prediction

From SHAP Scores to Feature Importance Scores

FiBiNet++: Reducing Model Size by Low Rank Feature Interaction Layer for CTR Prediction

A Guide to Feature Importance Methods for Scientific Inference

FeatNavigator: Automatic Feature Augmentation on Tabular Data

Enhancing Financial Market Predictions: Causality-Driven Feature Selection

Feature Interaction Fusion Self-Distillation Network For CTR Prediction

FLEN: Leveraging Field for Scalable CTR Prediction

Single Sample Feature Importance: An Interpretable Algorithm for Low-Level Feature Analysis

Ensembling improves stability and power of feature selection for deep learning models

Permutation-based identification of important biomarkers for complex diseases via machine learning models

FIARSE: Model-Heterogeneous Federated Learning via Importance-Aware Submodel Extraction

Predicting Scores of Various Aesthetic Attribute Sets by Learning from Overall Score Labels

Towards a More Reliable Interpretation of Machine Learning Outputs for Safety-Critical Systems using Feature Importance Fusion

LPFS: Learnable Polarizing Feature Selection for Click-Through Rate Prediction

Confident Feature Ranking

SMARTFEAT: Efficient Feature Construction through Feature-Level Foundation Model Interactions

Fusion Matters: Learning Fusion in Deep Click-through Rate Prediction Models