A boosted-trees method for name disambiguation
Author
Abstract
Suggested Citation
DOI: 10.1007/s11192-012-0681-1
Download full text from publisher
As the access to this document is restricted, you may want to search for a different version of it.
References listed on IDEAS
- Dag W. Aksnes, 2008. "When different persons have an identical author name. How frequent are homonyms?," Journal of the American Society for Information Science and Technology, Association for Information Science & Technology, vol. 59(5), pages 838-841, March.
- Natsuo Onodera & Mariko Iwasawa & Nobuyuki Midorikawa & Fuyuki Yoshikane & Kou Amano & Yutaka Ootani & Tadashi Kodama & Yasuhiko Kiyama & Hiroyuki Tsunoda & Shizuka Yamazaki, 2011. "A method for eliminating articles by homonymous authors from the large number of articles retrieved by author search," Journal of the American Society for Information Science and Technology, Association for Information Science & Technology, vol. 62(4), pages 677-690, April.
- Natsuo Onodera & Mariko Iwasawa & Nobuyuki Midorikawa & Fuyuki Yoshikane & Kou Amano & Yutaka Ootani & Tadashi Kodama & Yasuhiko Kiyama & Hiroyuki Tsunoda & Shizuka Yamazaki, 2011. "A method for eliminating articles by homonymous authors from the large number of articles retrieved by author search," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 62(4), pages 677-690, April.
- Ricardo G. Cota & Anderson A. Ferreira & Cristiano Nascimento & Marcos André Gonçalves & Alberto H. F. Laender, 2010. "An unsupervised heuristic-based hierarchical method for name disambiguation in bibliographic citations," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 61(9), pages 1853-1870, September.
- Vetle I. Torvik & Marc Weeber & Don R. Swanson & Neil R. Smalheiser, 2005. "A probabilistic similarity metric for Medline records: A model for author name disambiguation," Journal of the American Society for Information Science and Technology, Association for Information Science & Technology, vol. 56(2), pages 140-158, January.
- Alan L. Porter & Ismael Rafols, 2009. "Is science becoming more interdisciplinary? Measuring and mapping six research fields over time," Scientometrics, Springer;Akadémiai Kiadó, vol. 81(3), pages 719-745, December.
- Ciriaco Andrea D'Angelo & Cristiano Giuffrida & Giovanni Abramo, 2011. "A heuristic approach to author name disambiguation in bibliometrics databases for large-scale research assessments," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 62(2), pages 257-269, February.
- Steven Wooding & Kate Wilcox-Jay & Grant Lewison & Jonathan Grant, 2006. "Co-author inclusion: A novel recursive algorithmic method for dealingwith homonyms in bibliometric analysis," Scientometrics, Springer;Akadémiai Kiadó, vol. 66(1), pages 11-21, January.
- Li Tang & John P. Walsh, 2010. "Bibliometric fingerprints: name disambiguation based on approximate structure equivalence of cognitive maps," Scientometrics, Springer;Akadémiai Kiadó, vol. 84(3), pages 763-784, September.
- Ciriaco Andrea D'Angelo & Cristiano Giuffrida & Giovanni Abramo, 2011. "A heuristic approach to author name disambiguation in bibliometrics databases for large‐scale research assessments," Journal of the American Society for Information Science and Technology, Association for Information Science & Technology, vol. 62(2), pages 257-269, February.
Citations
Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
Cited by:
- Jinseok Kim & Jenna Kim, 2018. "The impact of imbalanced training data on machine learning for author name disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 117(1), pages 511-526, October.
- Wang, Jian, 2016. "Knowledge creation in collaboration networks: Effects of tie configuration," Research Policy, Elsevier, vol. 45(1), pages 68-80.
- Maxim Kotsemir & Sergey Shashnov, 2017. "Measuring, analysis and visualization of research capacity of university at the level of departments and staff members," Scientometrics, Springer;Akadémiai Kiadó, vol. 112(3), pages 1659-1689, September.
- Deyun Yin & Kazuyuki Motohashi & Jianwei Dang, 2020. "Large-scale name disambiguation of Chinese patent inventors (1985–2016)," Scientometrics, Springer;Akadémiai Kiadó, vol. 122(2), pages 765-790, February.
- Pascal Cuxac & Jean-Charles Lamirel & Valerie Bonvallot, 2013. "Efficient supervised and semi-supervised approaches for affiliations disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 97(1), pages 47-58, October.
- Dominik P. Heinisch & Guido Buenstorf, 2018. "The next generation (plus one): an analysis of doctoral students’ academic fecundity based on a novel approach to advisor identification," Scientometrics, Springer;Akadémiai Kiadó, vol. 117(1), pages 351-380, October.
- Wang, Jian & Hicks, Diana, 2015. "Scientific teams: Self-assembly, fluidness, and interdependence," Journal of Informetrics, Elsevier, vol. 9(1), pages 197-207.
- Jinseok Kim & Jinmo Kim & Jason Owen-Smith, 2019. "Generating automatically labeled data for author name disambiguation: an iterative clustering method," Scientometrics, Springer;Akadémiai Kiadó, vol. 118(1), pages 253-280, January.
- Fernanda Morillo & Ignacio Santabárbara & Javier Aparicio, 2013. "The automatic normalisation challenge: detailed addresses identification," Scientometrics, Springer;Akadémiai Kiadó, vol. 95(3), pages 953-966, June.
- Andrea Ancona & Roy Cerqueti & Gianluca Vagnani, 2023. "A novel methodology to disambiguate organization names: an application to EU Framework Programmes data," Scientometrics, Springer;Akadémiai Kiadó, vol. 128(8), pages 4447-4474, August.
- Rehs, Andreas, 2021. "A supervised machine learning approach to author disambiguation in the Web of Science," Journal of Informetrics, Elsevier, vol. 15(3).
- Humaira Waqas & Muhammad Abdul Qadir, 2021. "Multilayer heuristics based clustering framework (MHCF) for author name disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 126(9), pages 7637-7678, September.
- Helena Mihaljević & Lucía Santamaría, 2021. "Disambiguation of author entities in ADS using supervised learning and graph theory methods," Scientometrics, Springer;Akadémiai Kiadó, vol. 126(5), pages 3893-3917, May.
- Janaína Gomide & Hugo Kling & Daniel Figueiredo, 2017. "Name usage pattern in the synonym ambiguity problem in bibliographic data," Scientometrics, Springer;Akadémiai Kiadó, vol. 112(2), pages 747-766, August.
- YIN Deyun & MOTOHASHI Kazuyuki, 2018. "Inventor Name Disambiguation with Gradient Boosting Decision Tree and Inventor Mobility in China (1985-2016)," Discussion papers 18018, Research Institute of Economy, Trade and Industry (RIETI).
- Omar Hernando Avila-Poveda, 2014. "Technical report: the trend of author compound names and its implications for authorship identity identification," Scientometrics, Springer;Akadémiai Kiadó, vol. 101(1), pages 833-846, October.
- Jinseok Kim & Jenna Kim & Jason Owen‐Smith, 2021. "Ethnicity‐based name partitioning for author name disambiguation using supervised machine learning," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 72(8), pages 979-994, August.
- Jinseok Kim & Jenna Kim, 2020. "Effect of forename string on author name disambiguation," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 71(7), pages 839-855, July.
- Sandra Cristina Oliveira & Juliana Cobre & Danilo Florentino Pereira, 2021. "A measure of reliability for scientific co-authorship networks using fuzzy logic," Scientometrics, Springer;Akadémiai Kiadó, vol. 126(6), pages 4551-4563, June.
- KM. Pooja & Samrat Mondal & Joydeep Chandra, 2021. "Exploiting similarities across multiple dimensions for author name disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 126(9), pages 7525-7560, September.
- Song, Min & Kim, Erin Hea-Jin & Kim, Ha Jin, 2015. "Exploring author name disambiguation on PubMed-scale," Journal of Informetrics, Elsevier, vol. 9(4), pages 924-941.
- Wang, Jian, 2014. "Unpacking the Matthew effect in citations," Journal of Informetrics, Elsevier, vol. 8(2), pages 329-339.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Jan Schulz, 2016. "Using Monte Carlo simulations to assess the impact of author name disambiguation quality on different bibliometric analyses," Scientometrics, Springer;Akadémiai Kiadó, vol. 107(3), pages 1283-1298, June.
- Shuiqing Huang & Bo Yang & Sulan Yan & Ronald Rousseau, 2014. "Institution name disambiguation for research assessment," Scientometrics, Springer;Akadémiai Kiadó, vol. 99(3), pages 823-838, June.
- Milojević, Staša, 2013. "Accuracy of simple, initials-based methods for author name disambiguation," Journal of Informetrics, Elsevier, vol. 7(4), pages 767-773.
- Ciriaco Andrea D’Angelo & Nees Jan Eck, 2020. "Collecting large-scale publication data at the level of individual researchers: a practical proposal for author name disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 123(2), pages 883-907, May.
- Rehs, Andreas, 2021. "A supervised machine learning approach to author disambiguation in the Web of Science," Journal of Informetrics, Elsevier, vol. 15(3).
- Jinseok Kim & Jason Owen-Smith, 2021. "ORCID-linked labeled data for evaluating author name disambiguation at scale," Scientometrics, Springer;Akadémiai Kiadó, vol. 126(3), pages 2057-2083, March.
- Jinseok Kim & Jenna Kim, 2020. "Effect of forename string on author name disambiguation," Journal of the Association for Information Science & Technology, Association for Information Science & Technology, vol. 71(7), pages 839-855, July.
- Jiang Wu & Xiu-Hao Ding, 2013. "Author name disambiguation in scientific collaboration and mobility cases," Scientometrics, Springer;Akadémiai Kiadó, vol. 96(3), pages 683-697, September.
- Alison M. J. Buchan & Eva Jurczyk & Ruth Isserlin & Gary D. Bader, 2016. "Global neuroscience and mental health research: a bibliometrics case study," Scientometrics, Springer;Akadémiai Kiadó, vol. 109(1), pages 515-531, October.
- Jinseok Kim & Jinmo Kim & Jason Owen-Smith, 2019. "Generating automatically labeled data for author name disambiguation: an iterative clustering method," Scientometrics, Springer;Akadémiai Kiadó, vol. 118(1), pages 253-280, January.
- Cova, Tânia F.G.G. & Jarmelo, Susana & Formosinho, Sebastião J. & de Melo, J. Sérgio Seixas & Pais, Alberto A.C.C., 2015. "Unsupervised characterization of research institutions with task-force estimation," Journal of Informetrics, Elsevier, vol. 9(1), pages 59-68.
- Jinseok Kim, 2018. "Evaluating author name disambiguation for digital libraries: a case of DBLP," Scientometrics, Springer;Akadémiai Kiadó, vol. 116(3), pages 1867-1886, September.
- Mehmet Ali Abdulhayoglu & Bart Thijs, 2017. "Use of ResearchGate and Google CSE for author name disambiguation," Scientometrics, Springer;Akadémiai Kiadó, vol. 111(3), pages 1965-1985, June.
- Thomas Gurney & Edwin Horlings & Peter van den Besselaar, 2012. "Author disambiguation using multi-aspect similarity indicators," Scientometrics, Springer;Akadémiai Kiadó, vol. 91(2), pages 435-449, May.
- Abramo, Giovanni & D'Angelo, Ciriaco Andrea & Di Costa, Flavia, 2019. "Diversification versus specialization in scientific research: Which strategy pays off?," Technovation, Elsevier, vol. 82, pages 51-57.
- Cathelijn J F Waaijer & Benoît Macaluso & Cassidy R Sugimoto & Vincent Larivière, 2016. "Stability and Longevity in the Publication Careers of U.S. Doctorate Recipients," PLOS ONE, Public Library of Science, vol. 11(4), pages 1-15, April.
- Deyun Yin & Kazuyuki Motohashi & Jianwei Dang, 2020. "Large-scale name disambiguation of Chinese patent inventors (1985–2016)," Scientometrics, Springer;Akadémiai Kiadó, vol. 122(2), pages 765-790, February.
- Gianluca Fabiano & Andrea Marcellusi & Giampiero Favato, 2020. "Public–private contribution to biopharmaceutical discoveries: a bibliometric analysis of biomedical research in UK," Scientometrics, Springer;Akadémiai Kiadó, vol. 124(1), pages 153-168, July.
- Giovanni Abramo & Ciriaco Andrea D’Angelo, 2022. "Drivers of academic engagement in public–private research collaboration: an empirical study," The Journal of Technology Transfer, Springer, vol. 47(6), pages 1861-1884, December.
- Lutz Bornmann & Werner Marx, 2014. "How to evaluate individual researchers working in the natural and life sciences meaningfully? A proposal of methods based on percentiles of citations," Scientometrics, Springer;Akadémiai Kiadó, vol. 98(1), pages 487-509, January.
More about this item
Keywords
Name disambiguation; Common names; Classification tree; Boosted trees;All these keywords.
Statistics
Access and download statisticsCorrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:scient:v:93:y:2012:i:2:d:10.1007_s11192-012-0681-1. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.