Train dev test split sklearn

Curious Student

Code: Python

2021-03-10 21:06:28

from sklearn.model_selection import train_test_split

X = df.drop(['target'],axis=1).values   # independant features
y = df['target'].values					# dependant variable

# Choose your test size to split between training and testing sets:
X_train, X_test, y_train, y_test = train_test_split(X, y, test_size=0.25, random_state=42)

4

Mark Jacobs

Code: Python

2021-02-24 04:22:47

import numpy as np
from sklearn.model_selection import train_test_split

X, y = np.arange(10).reshape((5, 2)), range(5)

X_train, X_test, y_train, y_test = train_test_split(
    X, y, test_size=0.33, random_state=42)

X_train
# array([[4, 5],
#        [0, 1],
#        [6, 7]])

y_train
# [2, 0, 3]

X_test
# array([[2, 3],
#        [8, 9]])

y_test
# [1, 4]

1

ian5v

Code: Python

2021-03-10 21:07:39

X_train, X_test, y_train, y_test = train_test_split(X, y, test_size=0.33, random_state=42)

3

David Refaeli

Code: Python

2021-03-10 21:07:05

 X_train, X_test, y_train, y_test 
    = train_test_split(X, y, test_size=0.2, random_state=1)

 X_train, X_val, y_train, y_val 
    = train_test_split(X_train, y_train, test_size=0.25, random_state=1) # 0.25 x 0.8 = 0.2

2

jwestmoreland

Code: Python

2021-03-10 21:04:55

train, validate, test = np.split(df.sample(frac=1), [int(.6*len(df)), int(.8*len(df))])

2

zoli

Code: Python

2021-08-21 19:30:38

from sklearn.model_selection import train_test_split
X_train, X_test, y_train, y_test = train_test_split(x, y, test_size=0.33, random_state=42)
print(X_train.shape, X_test.shape, y_train.shape, y_test.shape)

0

Related

New to Communities?