You are browsing as a guest. Sign up (or log in) to start making projects!

49m 5s logged

I have made the attempt to start sorting and generating insights for the entire OpenAlex Database. I have developed code to create and sort CSVs for every single topic (of which there are about 4000). As it turns out, there are about 400 million papers with abstracts, which means every topic has about 100 thousand papers, just enough for effective analysis. The calculations should take about a month, but can be paused and continued, etc. I have implemented efficiencies to make the project feasible. I am excited to see what I learn.

Doing this I hope to:

Identify cross disciplinary topic gaps
Generate insights on the structure of the entire database through heighrarchical clustering
Identify per term reductions across fields to analyze confidence in sentiment.

0
3

Comments 0

No comments yet. Be the first!