remove duplicates dataframe python code example
Example 1: remove duplicates based on two columns in dataframe
df.drop_duplicates(['A','B'],keep= 'last')
Example 2: drop duplicates pandas first column
import pandas as pd
# making data frame from csv file
data = pd.read_csv("employees.csv")
# sorting by first name
data.sort_values("First Name", inplace = True)
# dropping ALL duplicte values
data.drop_duplicates(subset ="First Name",keep = False, inplace = True)
# displaying data
print(data)
Example 3: remove duplicate columns python dataframe
df = df.loc[:,~df.columns.duplicated()]