subhash-pujari/MicrosoftAcademicSearchCrawler
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
This project is part of analysing the Twitter data to find the activity of computer scientist on Twitter. The steps followed are as: (i) parsing the dblp data (ii) filtering Twitter network for identified user in DBLP (iii) Find the specific field in which a researcher work based on conference he has published to (iv) Creating network of information flow (v) finding authors activity in the network (vi) how author interact in pairwise and some utility fucntion to compute the common tasks
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
This is the crawler for querying the microsoft academic search APIs in BFS(breadth first search way starting from the root node). We get a JSON response which is parsed and saved to the database.
This is the Lucene indexer i wrote to index the title of the publications.
This is a simple lucene search server that search through the lucene indexed directory. We have a simple http server in Java which listens for the incoming request and pass it to the lucene engine which in turn determine the data to be retrieved. The data is then converted to xml and send back as http reponse.
Public repository.