I have a dataframe like this (assuming one column):
column
[A,C,B,A]
[HELLO,HELLO,ha]
[test/1, test/1, test2]
The type of the column above is: dtype('O')
I would like to remove the duplicates here, resulting in:
column
[A,C,B] # - A
[HELLO, ha] # removing 1 hello
[test/1, test2] # removing 1 test/1
Then, I would like to sort the data
column
[A,B,C]
[ha, HELLO]
[test2, test/1] # assuming that number comes before /
I am struggling getting this done in a proper way. Hope anyone has nice ideas (would it make sense to transform to small lists?)