GraphDBLP: a system for analysing networks of computer scientists through graph databases: GraphDBLP
Autor: | Mirko Cesarini, Mario Mezzanzanica, Antonio Picariello, Fabio Mercorio, Vincenzo Moscato |
---|---|
Přispěvatelé: | Mezzanzanica, M, Mercorio, F, Cesarini, M, Moscato, V, Picariello, A |
Jazyk: | angličtina |
Rok vydání: | 2018 |
Předmět: |
Source code
Word embedding social network analysis Computer Networks and Communications Computer science media_common.quotation_subject Social network analysi graph database 02 engineering and technology computer.software_genre Knowledge extraction 020204 information systems 0202 electrical engineering electronic engineering information engineering Media Technology Semantic analytics Word2vec Social network analysis media_common Graph database Information retrieval Semantic Analytic Graph Hardware and Architecture Graph (abstract data type) 020201 artificial intelligence & image processing computer Software |
Popis: | This paper presents GraphDBLP, a system that models the DBLP bibliography as a graph database for performing graph-based queries and social network analyses. GraphDBLP also enriches the DBLP data through semantic keyword similarities computed via word-embedding. In this paper, we discuss how the system was formalized as a multi-graph, and how similarity relations were identified through word2vec. We also provide three meaningful queries for exploring the DBLP community to (i) investigate author profiles by analysing their publication records; (ii) identify the most prolific authors on a given topic, and (iii) perform social network analyses over the whole community. To date, GraphDBLP contains 5+ million nodes and 24+ million relationships, enabling users to explore the DBLP data by referencing more than 3.3 million publications, 1.7 million authors, and more than 5 thousand publication venues. Through the use of word-embedding, more than 7.5 thousand keywords and related similarity values were collected. GraphDBLP was implemented on top of the Neo4j graph database. The whole dataset and the source code are publicly available to foster the improvement of GraphDBLP in the whole computer science community. |
Databáze: | OpenAIRE |
Externí odkaz: |