Features for Author Disambiguation

저자 식별을 위한 자질 비교

  • 강인수 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 이승우 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 정한민 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 김평 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 구희관 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 이미경 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 성원경 (한국과학기술정보연구원 정보서비스연구팀) ;
  • 박동인 (한국과학기술정보연구원 정보서비스연구팀)
  • Published : 2008.02.28


There exists a many-to-many mapping relationship between persons and their names. A person may have multiple names, and different persons may share the same name. These synonymous and homonymous names may severely deteriorate the recall and precision of the person search, respectively. This study addresses the characteristics of features for resolving homonymous author names appearing in citation data. As disambiguation features, previous works have employed citation-internal features such as co-authorship, titles of articles, titles of publications as well as citation-external features such as emails, affiliations, Web evidences. To the best of our knowledge, however, there has been no literature to deal with the influences of features on author disambiguation. This study analyzes the effect of individual features on author resolution using a large-scale test set for Korean.


