
I used StreamSets in one of my ETL project where we were working on Apache Spark for handling bulk data. The best thing I feel is it provides an easy way to configure pipelines so that we were able to process tasks easily using the same. Also apart from this it also helped in monitoring the same. Even for beginners it is very easy to learn. Review collected by and hosted on G2.com.
While doing performance testing it took a lot of time for around 8-10 million records. That could be improved further. Also I think there is a scope to improve the details of errors in a more detailed way. Review collected by and hosted on G2.com.