Pull out all the stops: Textual analysis via punctuation sequences

  • 2020-01-16 17:02:45
  • Alexandra N. M. Darmon, Marya Bazzi, Sam D. Howison, Mason A. Porter
  • 0

Abstract

Whether enjoying the lucid prose of a favorite author or slogging throughsome other writer's cumbersome, heavy-set prattle (full of parentheses, emdashes, compound adjectives, and Oxford commas), readers will notice stylisticsignatures not only in word choice and grammar, but also in punctuation itself.Indeed, visual sequences of punctuation from different authors producemarvelously different (and visually striking) sequences. Punctuation is alargely overlooked stylistic feature in "stylometry", the quantitative analysisof written text. In this paper, we examine punctuation sequences in a corpus ofliterary documents and ask the following questions: Are the properties of suchsequences a distinctive feature of different authors? Is it possible todistinguish literary genres based on their punctuation sequences? Do thepunctuation styles of authors evolve over time? Are we on to somethinginteresting in trying to do stylometry without words, or are we full of soundand fury (signifying nothing)?

 

Quick Read (beta)

loading the full paper ...