stylo. R package for stylometric analyses
223100_english_novels. A benchmark corpus of 100 English novels, covering the 19th and the beginning of the 20th century
24stylo_howto. Documentation for 'stylo', an R package for text analysis, suitable for authorship attribution, stylometry, and other multivariate analysis tasks in the domain of (literary) texts
13A_Small_Collection_of_British_Fiction. A selection of 28 classic British novels from the 19th century (including a few late 18th-century items). Full text versions, in plain text format, harvested from trustworthy public domain sites.
11100_polish_novels. A benchmark corpus of 100 Polish novels, covering the 19th and the beginning of the 20th century
10tidystopwords. Customizable lists of stopwords in multiple languages
668_german_novels. A benchmark corpus of 68 German novels, covering the 19th and the beginning of the 20th century
6beyond_Manhattan. Data and code supporting the study Manhattan, Euclidean, and their Siblings: Exploring Exotic Similarity Measures in Text Classification
4DHAbstracts_biblio_style. A bibliographic style definition for Digital Humanities 2016 conference
4NT_Vulgate.
3computationalstylistics.github.io. SCSS
3presentations. HTML
3stylometry_of_papyri.
2HVEE.00.016. Materials and slides for the course Data Science and Digital Humanities
2HVEE.00.046. Materials and slides for the course Introduction to Digital Humanities
1history_of_words. JavaScript
1topic_modeling_and_NLP.
1litRiddle. The package contains the data of a reader survey about fiction in Dutch, a description of the novels the readers rated, and the results of stylistic measurements of the novels. The package also contains functions to combine, analyze, and visualize these data.
1preprints. A selection of pre-prints by the members of the Group
1word_frequencies. Code for the study on improving relative word frequencies
1RdlR_for_rolling_classify. Roman de la Rose
1