Skip to main content
Article
A Syntactic Characterization of Authorship Style Surrounding Proper Names
Digital Scholarship in the Humanities (2015)
  • Ana Lucic, DePaul University
  • Catherine Blake, University of Illinois at Urbana-Champaign
Abstract
Accurately determining who wrote a manuscript has captivated scholars of literary history for centuries, as the true author can have important ramifications in religion, law, literary studies, philosophy, and education. A wide array of lexical, character, syntactic, semantic, and application-specific features have been proposed to represent a text so that authorship attribution can be established automatically. Although surface-level features have been tested extensively, few studies have systematically explored high-level features, in part due to limitations in the natural language processing techniques required to capture high-level features. However, high-level features, such as sentence structure, are used subconsciously by a writer and thus may be more consistent than surface-level features, such as word choice. In this article, we introduce a new high-level feature based on local syntactic dependencies that an author uses when referring to a named entity (in our case a person’s name). The series of experiments in the contexts of movie reviews reveal how the amount of data in both the training and test sets influences predictive performance. Finally, we measure authorship consistency with respect to this new feature and show how consistency influences predictive performance. These results provide other researchers with a new model for how to evaluate new features and suggest that the local syntactic dependencies warrant further investigation.
Keywords
  • authorship attribution,
  • stylometry,
  • proper names
Disciplines
Publication Date
2015
DOI
10.1093/llc/fqt033
Citation Information
Lucic, A. & Blake, L. C. (2015). A Syntactic Characterization of Authorship Style Surrounding Proper Names. Digital Scholarship in the Humanities Advance Access Published June 29, 2013. doi:10.1093/llc/fqt033.
Creative Commons license
Creative Commons License
This work is licensed under a Creative Commons CC_BY International License.