ScrapPaper: A web scrapping method to extract journal information from PubMed and Google Scholar search result using Python.
Rafsanjani, M. R.
Show abstract
This paper introduces a program called ScrapPaper, a simple Python script that use web-scraping method to extract journal information from PubMed and Google Scholar search results page. Currently the motivation behind the program development is trying to solve a problem to obtain scientific literatures information especially the title and link and save as a list for further use such as in meta-analysis and comparative study of literatures. ScrapPaper advantage that it is simple to use with no prior programming experience and get the results ready within minutes (depending on the total search result). Web-scrapping is a very powerful method to extract information from the web and ScrapPaper employ several server friendly approaches accessing both PubMed and Google Scholar site. O_FIG_DISPLAY_L [Figure 1] M_FIG_DISPLAY C_FIG_DISPLAY
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.