subhash-pujari/MicrosoftAcademicSearchCrawler
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
Example code from my PyCon APAC 2012 talk.
Public repository record indexed from GitHub. Explore verified star velocity metrics, source code repositories, and curated developer tool directories across the GitHubRepo ecosystem.
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
This is the Lucene indexer i wrote to index the title of the publications.
This is a simple lucene search server that search through the lucene indexed directory. We have a simple http server in Java which listens for the incoming request and pass it to the lucene engine which in turn determine the data to be retrieved. The data is then converted to xml and send back as http reponse.
Public repository.