Repository Issues

mGalarnyk/DSE230_Data_Analysis_Using_Hadoop_and_Spark_UCSD

Map-reduce, streaming analysis, and external memory algorithms and their implementation using the Hadoop and its eco-system: HBase, Hive, Pig and Spark. The class will include assignment of analyzing large existing databases.

View on GitHub
Stars
 (34 stars)
Forks
 (21 forks)
Indexed issues
 (0 indexed issues)
open beginner issues
 (0 open beginner issues)
Latest indexed
Aug 17, 2026
Last GitHub push
Apr 3, 2017
License
No license data
Contributing guide
No contributing guide
Code of conduct
No code of conduct
Dominant language
Jupyter Notebook
PR merge metrics
 (No merged PRs in 30d)
Beginner labels
No beginner labels indexed

Issues

0 open indexed issues

No open indexed issues found for this repository.