On the use of topological features and hierarchical characterization for disambiguating names in collaborative networks

Amancio, Diego R.; Oliveira Jr., Osvaldo N.; Costa, Luciano da F.

doi:10.1209/0295-5075/99/48002

Physics > Physics and Society

arXiv:1302.4504 (physics)

[Submitted on 19 Feb 2013]

Title:On the use of topological features and hierarchical characterization for disambiguating names in collaborative networks

Authors:Diego R. Amancio, Osvaldo N. Oliveira Jr., Luciano da F. Costa

View PDF

Abstract:Many features of complex systems can now be unveiled by applying statistical physics methods to treat them as social networks. The power of the analysis may be limited, however, by the presence of ambiguity in names, e.g., caused by homonymy in collaborative networks. In this paper we show that the ability to distinguish between homonymous authors is enhanced when longer-distance connections are considered, rather than looking at only the immediate neighbors of a node in the collaborative network. Optimized results were obtained upon using the 3rd hierarchy in connections. Furthermore, reasonable distinction among authors could also be achieved upon using pattern recognition strategies for the data generated from the topology of the collaborative network. These results were obtained with a network from papers in the arXiv repository, into which homonymy was deliberately introduced to test the methods with a controlled, reliable dataset. In all cases, several methods of supervised and unsupervised machine learning were used, leading to the same overall results. The suitability of using deeper hierarchies and network topology was confirmed with a real database of movie actors, with the additional finding that the distinguishing ability can be further enhanced by combining topology features and long-range connections in the collaborative network.

Subjects:	Physics and Society (physics.soc-ph); Digital Libraries (cs.DL); Information Retrieval (cs.IR); Social and Information Networks (cs.SI)
Cite as:	arXiv:1302.4504 [physics.soc-ph]
	(or arXiv:1302.4504v1 [physics.soc-ph] for this version)
	https://doi.org/10.48550/arXiv.1302.4504
Journal reference:	Europhysics Letters (2012) 99 48002
Related DOI:	https://doi.org/10.1209/0295-5075/99/48002

Submission history

From: Diego Amancio Raphael [view email]
[v1] Tue, 19 Feb 2013 02:00:01 UTC (1,670 KB)

Physics > Physics and Society

Title:On the use of topological features and hierarchical characterization for disambiguating names in collaborative networks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Physics > Physics and Society

Title:On the use of topological features and hierarchical characterization for disambiguating names in collaborative networks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators