DNN2LR: Automatic Feature Crossing for Credit Scoring

02/24/2021
by   Qiang Liu, et al.
4

Credit scoring is a major application of machine learning for financial institutions to decide whether to approve or reject a credit loan. For sake of reliability, it is necessary for credit scoring models to be both accurate and globally interpretable. Simple classifiers, e.g., Logistic Regression (LR), are white-box models, but not powerful enough to model complex nonlinear interactions among features. Fortunately, automatic feature crossing is a promising way to find cross features to make simple classifiers to be more accurate without heavy handcrafted feature engineering. However, credit scoring is usually based on different aspects of users, and the data usually contains hundreds of feature fields. This makes existing automatic feature crossing methods not efficient for credit scoring. In this work, we find local piece-wise interpretations in Deep Neural Networks (DNNs) of a specific feature are usually inconsistent in different samples, which is caused by feature interactions in the hidden layers. Accordingly, we can design an automatic feature crossing method to find feature interactions in DNN, and use them as cross features in LR. We give definition of the interpretation inconsistency in DNN, based on which a novel feature crossing method for credit scoring prediction called DNN2LR is proposed. Apparently, the final model, i.e., a LR model empowered with cross features, generated by DNN2LR is a white-box model. Extensive experiments have been conducted on both public and business datasets from real-world credit scoring applications. Experimental shows that, DNN2LR can outperform the DNN model, as well as several feature crossing methods. Moreover, comparing with the state-of-the-art feature crossing methods, i.e., AutoCross, DNN2LR can accelerate the speed for feature crossing by about 10 to 40 times on datasets with large numbers of feature fields.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/22/2020

DNN2LR: Interpretation-inspired Feature Crossing for Real-world Tabular Data

For sake of reliability, it is necessary for models in real-world applic...
research
05/22/2021

Explainable Enterprise Credit Rating via Deep Feature Crossing Network

Due to the powerful learning ability on high-rank and non-linear feature...
research
04/29/2019

AutoCross: Automatic Feature Crossing for Tabular Data in Real-World Applications

Feature crossing captures interactions among categorical features and is...
research
12/03/2020

Explainable AI for Interpretable Credit Scoring

With the ever-growing achievements in Artificial Intelligence (AI) and t...
research
08/17/2017

Deep & Cross Network for Ad Click Predictions

Feature engineering has been the key to the success of many prediction m...
research
01/29/2020

Credit Scoring for Good: Enhancing Financial Inclusion with Smartphone-Based Microlending

Globally, two billion people and more than half of the poorest adults do...
research
09/21/2022

Monotonic Neural Additive Models: Pursuing Regulated Machine Learning Models for Credit Scoring

The forecasting of credit default risk has been an active research field...

Please sign up or login with your details

Forgot password? Click here to reset