Binarization in Spark and Scala

Here is my 100 line ( or less ) – of Spark with Scala code – to perform Binarization of text file

This is a very simplistic solution – which was written in 1-2 hours of heads down coding time.

I had previously coded the same algorithm in MR ( Java Code – with multiple stages of MR ) and it took more than 400 lines of coding.

Advertisements

Leave a Reply

Fill in your details below or click an icon to log in:

WordPress.com Logo

You are commenting using your WordPress.com account. Log Out / Change )

Twitter picture

You are commenting using your Twitter account. Log Out / Change )

Facebook photo

You are commenting using your Facebook account. Log Out / Change )

Google+ photo

You are commenting using your Google+ account. Log Out / Change )

Connecting to %s