train_size parameter in train_test_split represents the proportion of the dataset to include in the train split code example

Example 1: sklearn split train test

import numpy as np
from sklearn.model_selection import train_test_split

X, y = np.arange(10).reshape((5, 2)), range(5)

X_train, X_test, y_train, y_test = train_test_split(
    X, y, test_size=0.33, random_state=42)

# array([[4, 5],
#        [0, 1],
#        [6, 7]])

# [2, 0, 3]

# array([[2, 3],
#        [8, 9]])

# [1, 4]

Example 2: train_size

You have to specify this parameter only if you’re not specifying the test_size. This is the same as test_size, but instead you tell the class what percent of the dataset you want to split as the training set.