I have a data frame that looks like this:
| Image | Similar Images |
| ------| -------------- |
| 1 | [1, 2, 6] |
| 2 | [2, 1, 6] |
| 3 | [3, 4] |
| 4 | [4, 3] |
| 5 | [5] |
| 6 | [6, 1, 2] |
And I want to make clusters of similar images and label them. What I aim to do would look something like this:
| Image | Similar Images | Label |
| ------| -------------- |-------|
| 1 | [1, 2, 6] | 1 |
| 2 | [2, 1, 6] | 1 |
| 3 | [3, 4] | 2 |
| 4 | [4, 3] | 2 |
| 5 | [5] | 3 |
| 6 | [6, 1, 2] | 1 |
Is there an efficient way to do this? I have limited computing resources and around 178000 images, which is why I'm wondering if there are any efficient existing methods or packages that could perform (part of) this task.