I would like to do some DBSCAN on Spark. I have currently found 2 implementations:
I have tested the first one with the sbt configuration given in its github but:
functions in the jar are not the same as those in the doc or in the source on github. For example, I cannot find the train function in the jar
I manage to run a test with the fit function (found in the jar) but a bad configuration of epsilon (a little to big) put the code in an infinite loop.
code :
val model = DBSCAN.fit(eps, minPoints, values, parallelism)
Has someone managed to do someting with the first library?
Has someone tested the second one?