Pull out all the stops: Textual analysis via punctuation sequences

  • 2018-12-31 18:48:20
  • Alexandra N. M. Darmon, Marya Bazzi, Sam D. Howison, Mason A. Porter
  • 25

Abstract

Whether enjoying the lucid prose of a favorite author or slogging throughsome other writer's cumbersome, heavy-set prattle (full of parentheses,em-dashes, compound adjectives, and Oxford commas), readers will noticestylistic signatures not only in word choice and grammar, but also inpunctuation itself. Indeed, visual sequences of punctuation from differentauthors produce marvelously different (and visually striking) sequences.Punctuation is a largely overlooked stylistic feature in "stylometry'', thequantitative analysis of written text. In this paper, we examine punctuationsequences in a corpus of literary documents and ask the following questions:Are the properties of such sequences a distinctive feature of differentauthors? Is it possible to distinguish literary genres based on theirpunctuation sequences? Do the punctuation styles of authors evolve over time?Are we on to something interesting in trying to do stylometry without words, orare we full of sound and fury (signifying nothing)?

 

Introduction (beta)

None

 

Conclusion (beta)

None