jiangyu/SparkLDA
implement LDA algorithm using spark
This work aims to re-engineer the Hadoop Distributed File System (HDFS) so that it can be 1) highly available, and 2) horizontally scalable. This is achieved by replacing the central master server with a distributed real-time database (in our implementation, MySQL Cluster).
Public repository record indexed from GitHub. Explore verified star velocity metrics, source code repositories, and curated developer tool directories across the GitHubRepo ecosystem.
implement LDA algorithm using spark
Detect hadoop log from aggregation logs using spark
Public repository.
Public repository.