GraphDBLP: a system for analysing networks of computer scientists through graph databases: GraphDBLP

Autor: Mirko Cesarini, Mario Mezzanzanica, Antonio Picariello, Fabio Mercorio, Vincenzo Moscato
Přispěvatelé: Mezzanzanica, M, Mercorio, F, Cesarini, M, Moscato, V, Picariello, A
Jazyk: angličtina
Rok vydání: 2018
Předmět:
Popis: This paper presents GraphDBLP, a system that models the DBLP bibliography as a graph database for performing graph-based queries and social network analyses. GraphDBLP also enriches the DBLP data through semantic keyword similarities computed via word-embedding. In this paper, we discuss how the system was formalized as a multi-graph, and how similarity relations were identified through word2vec. We also provide three meaningful queries for exploring the DBLP community to (i) investigate author profiles by analysing their publication records; (ii) identify the most prolific authors on a given topic, and (iii) perform social network analyses over the whole community. To date, GraphDBLP contains 5+ million nodes and 24+ million relationships, enabling users to explore the DBLP data by referencing more than 3.3 million publications, 1.7 million authors, and more than 5 thousand publication venues. Through the use of word-embedding, more than 7.5 thousand keywords and related similarity values were collected. GraphDBLP was implemented on top of the Neo4j graph database. The whole dataset and the source code are publicly available to foster the improvement of GraphDBLP in the whole computer science community.
Databáze: OpenAIRE