Ben68952
0
Q:

scikit learn to identify highly correlated features

# Create correlation matrix
corr_matrix = df.corr().abs()

# Select upper triangle of correlation matrix
upper = corr_matrix.where(np.triu(np.ones(corr_matrix.shape), k=1).astype(np.bool))

# Find index of feature columns with correlation greater than 0.95
to_drop = [column for column in upper.columns if any(upper[column] > 0.95)]
0
df[df.columns[1:]].corr()['LoanAmount'][:]
0

New to Communities?

Join the community