remove duplicates from dataframe python code example
Example 1: python: remove duplicate in a specific column
df = df.drop_duplicates(subset=['Column1', 'Column2'], keep='first')
Example 2: remove duplicates based on two columns in dataframe
df.drop_duplicates(['A','B'],keep= 'last')
Example 3: drop duplicates pandas first column
import pandas as pd
data = pd.read_csv("employees.csv")
data.sort_values("First Name", inplace = True)
data.drop_duplicates(subset ="First Name",keep = False, inplace = True)
print(data)
Example 4: remove duplicate row in df
df = df.drop_duplicates()
Example 5: remove duplicate columns python dataframe
df = df.loc[:,~df.columns.duplicated()]
Example 6: df remove duplicate rows
df = df.drop_duplicates()
p