osDQ dedicated to create apache spark based data pipeline using JSON
...This uses java API of apache spark. It can run in local mode also.
Get json example at https://github.com/arrahtech/osdq-spark
How to run
Unzip the zip file
Windows : java -cp .\lib\*;osdq-spark-0.0.1.jar org.arrah.framework.spark.run.TransformRunner -c .\example\samplerun.json
Mac UNIX
java -cp ./lib/*:./osdq-spark-0.0.1.jar org.arrah.framework.spark.run.TransformRunner -c ./example/samplerun.json
For those on windows, you need to have hadoop distribtion unzipped on local drive and HADOOP_HOME set. Also copy winutils.exe from here into HADOOP_HOME\bin
e-lib is a simple library management software for GNU/Linux built using Python+Postgres+Qt+Apache. Suiting the Engineering Profession, we can keep/manage records of books, periodicals, articles, funds as well as various other academic activities/events.