I have defined a Pandas dataframe which contains a column name as 'emp_title'. I want to see the total number of unique entries in that specific column.
I used:
len(df['emp_title'].unique())
which gives me a value of 173106
whereas when I use:
df['emp_title'].nunique()
it gives me a value of 173105 which should be the actual size.
Can any one explain why I shouldn't be using the code with the len() function. Or is there probably an issue with the dataset at play here?