subhash-pujari/MicrosoftAcademicSearchCrawler
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
Discovered public repositories for subhash-pujari in the GitHub catalog.
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
This is the Lucene indexer i wrote to index the title of the publications.
This is a simple lucene search server that search through the lucene indexed directory. We have a simple http server in Java which listens for the incoming request and pass it to the lucene engine which in turn determine the data to be retrieved. The data is then converted to xml and send back as http reponse.
Public repository.
This project is part of analysing the Twitter data to find the activity of computer scientist on Twitter. The steps followed are as: (i) parsing the dblp data (ii) filtering Twitter network for identified user in DBLP (iii) Find the specific field in which a researcher work based on conference he has published to (iv) Creating network of information flow (v) finding authors activity in the network (vi) how author interact in pairwise and some utility fucntion to compute the common tasks
This web app is to recommend the publication based on the current publication provided.
python apis to extract the data from dblp and insert datastructure
Wikipedia scrapper to get compter science conference list.
This is a python scraper library to extract links from google search.
Android application that act as cache for the phone. A user can add his file into the application, and then application take care of transfering the files to the remote system in case there is not enough space on the phone. To test the prototype we used Apache tomcat with a servlet to listen for the get and post request from the application. In case of not enough space in the mobile the cache replacement logic is executed to find out the files to be deleted from the server.
Public repository.