This is a pretty interesting benchmark study, although the headline is a bit misleading because Hadoop isn’t really optimized for graph analysis. When you look at comparisons to Spark, GraphLab and other platforms, it seems the decision of what to choose might come down to data volume, acceptable latency and cost, especially when considered against the value of that graph workload. Projects like Giraph and other YARN-enabled engines might make Hadoop look better, too.

You’re subscribed! If you like, you can update your settings

Story posted at: bigdatarepublic.com

Comments have been disabled for this post