turian/batch-elki-cluster
Public repository.
Discovered public repositories for turian in the GitHub catalog.
Public repository.
Find the best model, using random hyperparameter optimization, using scikit-learn
Public repository.
Public repository.
The nanoc web site
ASKBOT is a StackOverflow-like Q&A forum, based on CNPROG.
Compress short strings, using the Huffman algorithm.
django-lazysignup is a package designed to allow users to interact with a site as if they were authenticated users, but without signing up. At any time, they can convert their temporary user account to a real user account.
fabric recipes, primarily for deploying Ubuntu and EC2 instances.
Postprocess XML output from wikiprep (Wikipedia preprocessor) into JSON
ESA implementation using Wikiprep output
Grab all Wikipedia abstracts, in all languages
Updates to Zope's keyphrase extractor (forked from 1.1.0)
scikit-learn main repo
Dmoz RDF parser
Recipes for scikits.learn
A permissive version of the Python SimpleXMLRPCServer, which can correct errant XML input from the client.
Java implementation of the TextRank algorithm by Mihalcea, et al. http://lit.csci.unt.edu/index.php/Graph-based_NLP
Extension for Mozilla Firefox and Google Chrome to save all of your open tabs to a text file (window/tab index, URL and title of each tab)
Fork of Xavier's code, for sparse sampling reconstruction [Theano based deep ANN learning code]
Javascript autocomplete, with MySQL/PHP backend
Deploy FatFree CRM on EC2
Firefox extension to select all workers in vWorker search results page
XML-RPC version of the Stanford POS tagger
Python code for accessing the CrowdFlower API
Perform a biased sample of text data
KEA 5.0 (keyphrase extraction software), modified to be an XML-RPC service
jsMath support for OSQA
Random projection library for Python, converting a dictionary to low-dimensional numpy matrix
IM-like application for Pinax social networks (Django), that allow you to see which friends are online and chat them
OSQA branch, with some fixes
Induce word representations using random indexing (RI)
Didactic example of information retrieval, computing the similarity of two twitter users
Extract faces from video clips; generate training data for pose-invariant face features
Preprocess text for NLP (tokenizing, lowercasing, stemming, sentence splitting, etc.)
Install OSQA on webfaction
Implementation of neural language models, in particular Collobert + Weston (2008) and a stochastic margin-based version of Mnih's LBL.
KDDCup 2005 query classification with word representations
KDDCup 2005 query classification with word representations
Train a CRF for syntactic chunking (CoNLL2000), and use word representations
A static file blog engine/compiler, inspired by jekyll
Python methods to interact with the Crunchbase API v1.
In Python, read the .80 file format, for 80legs web crawl results.
A neural language model, intended to produce embeddings for a linear classifier
An article about scientific collaboration
HMM model for word representations, using the method of Huang + Yates (2009).
A bliki (blog+wiki) compiler, inspired by ikiwiki
Common scripts, mainly for text processing and experimental control
Common Python library, especially for text processing and controlling experimental runs
A neural network with a sparse input, for predicting decisions of a natural language syntax parser.