joeywen/trident-memcached
Trident state implementation for Memcached
Discovered public repositories for joeywen in the GitHub catalog.
Trident state implementation for Memcached
mt code
HBase Writes Distributor
Trident-ML : A realtime online machine learning library
This project implements online-k-clustering algorithm as mentioned in this paper(http://cseweb.ucsd.edu/~dasgupta/291/lec6.pdf). It produces REALTIME k-clustering on an infinite stream of data. It is implemented on top of twitter storm and uses cassandra as database. It deals with 2-dimensional matrices and clusters in Euclidean space.
hadoop (hadoop,hive,hue,hbase) deployer
aka "Bayesian Methods for Hackers": An introduction to Bayesian methods + probabilistic programming with a computation/understanding-first, mathematics-second point of view. All in pure Python ;)
Public repository.
利用Netty写的一个简单的NIO Server
利用java nio 包写的一个简单的Java server
能够直接应用于Solr的SVMRank java代码。
针对中文,演示Markdown的各种语法
Zarkov is a Lightweight Map-Reduce Framework
Spark In MapReduce (SIMR)
MapReduce course at the University of Maryland for Spring 2013
Data-Intensive Text Processing with MapReduce
PageRank, Spark version.
Most queries issued to a search engine are ambiguous at some level. This presents the search engine with a dilemma. On the one hand, if it focuses on the most likely interpretation of the query, it does not provide any utility to users that had a different intent. On the other hand, if it tries to provides at least one relevant results for all query intents, it will not cover any intent particularly well. Dynamic Ranking is a retrieval model that combines these otherwise contradictory goals of result diversification and high coverage. The key idea is to make the result ranking dynamic, allowing limited and well-controllable change based on interactive user feedback. Unlike conventional rankings that remain static after the query was issued, a dynamic ranking allows and anticipates user activity.
Public repository.
A simple mark-sweep garbage collector in Java
some algorithms implemented by mapreduce
A Python package for optimization
skiplist implemented by c++
A simple java mongodb client
复旦的中文自然语言工具包
本工程只是做些实验性的工作。利用JNI,为Spark-0.6.0开发了一点点C++ API,目前能成功运行wordcount,kmeans,还在开发中,一些环境变量的配置没有更改,如有想下载运行的话,可能要花点时间配置一下。以后还会继续更新,等到一些配置好了,会更新通知。e-mail:zhgwen@outlook.com