SongSim. React web app for drawing self-similarity matrices from text
553char-rbm. Character-level RBMs for short text
111pejorative-compounds. Analysing patterns in English noun-noun pejorative compounds on Reddit
110reddit-dubious-spelling. Analyzing words Redditors aren't sure how to spell
47lalala. Investigating repetitiveness in pop music using compression algorithms
43instacart-basket-prediction. RNNs (and other stuff) for Kaggle's Instacart order prediction competition
27wiki-pageview-floor. Finding and analysing the least viewed articles on English Wikipedia
26tour-of-heroes. some crummy incremental game made with angular 2
24atypicality. Jupyter Notebook
19reponames-dataset. Dataset of 4.6m GitHub repository names
16sketch-rnn-experiments. Experiments with Sketch-RNN and the Quick Draw dataset
11crawl-coroner. Analyzing completed games ('morgue files') of Dungeon Crawl Stone Soup
11favicon-scraper. Scraping a dataset of favicons.
7colinmorris.github.com. HTML
7lm1b-notebook. Various scripts used while playing around with Google Brain's billion word language model
7song-repetition. Are Pop Lyrics Getting More Repetitive?
4snl-notebooks. Some ipython notebooks analyzing SNL data
4csc236. Course website for CSC236 winter 2020
4reddit-username-suffixes. Little one-off experiment analysing the frequency of numerical suffixes in reddit usernames
4circles-of-hell-ngrams. An ngram experiment
3reddit-misspelling-trends. Trends in the frequency of common misspellings on Reddit over time
3size-of-an-x. Looking at the changing frequencies of different size analogies in Google Books ngram data.
2Tone.js. A Web Audio framework for making interactive music in the browser.
2unique-country-prefixes. Infographic ranking countries by the length of the shortest unique prefix that identifies them
2moz-graphs. Python
2testify. Unit test generator proof of concept (intended to be used in Exercism)
1openai-gym-sandbox. Experiments with OpenAI gym environments (https://gym.openai.com/)
1lm-sentences. Web app for visualizing how a language model assigns probabilities to sentences
1wiki-controversial-titles. Analyzing Wikipedia articles with the most debated titles
1xscreensaver-text-toys. Shell
1wiki-listicle-generation. Semi-automatically writing Wikipedia list articles using Wikidata/Wikipedia APIs
1tldr. :books: Simplified and community-driven man pages
1rm_stats. Web app for visualizing rename discussions on Wikipedia
1