This is your work, valued
Machine Learning researcher focused on Pragmatic AI alignment. Based in Perth, Australia. Waiting to be uploaded.
rl-portfolio-management. Attempting to replicate "A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem" https://arxiv.org/abs/1706.10059 (and an openai gym environment)
562keywordshitter2. A website to find long-tail keywords using search suggestions
221viz_torch_optim. Videos of deep learning optimizers moving on 3D problem-landscapes
107attentive-neural-processes. implementing "recurrent attentive neural processes" to forecast power usage (w. LSTM baseline, MCDropout)
100world-models-sonic-pytorch. Attempt at reinforcement learning with curiosity for Sonic the Hedgehog games. Number 149 on OpenAI retro contest leaderboard, but more work needed
33satellite_leak_detection. Detect water leaks from satellite images using machine learning
31thinkpad-x380-yoga-scripts. ThinkPad x380 scripts for GNU/Linux that allow Yoga features (tablet mode, pen mode, auto rotate, auto brightness)
23pipe-segmentation. Using machine learning to find water pipelines in aerial images with 73% accuracy
22phoneme2grapheme. Teaching machines to spell with deep learning (acc=>80%) e.g. a model hears "pɹˈaʊd˺ɚ" and writes "prowder" (but it should be "prouder")
19prob_jsonformer. Generate Structured JSON with probs from Language Models
17quiet-star. investigate Quiet-STaR paper, and it's thought scratchpad
14awesome-satellite-imagery-competitions. List of machine learning competitions for satellite imagery and remote sensing.
11pytorch-pretrained-BERT_horror-genor. PyTorch BERT model - for generating horror
11rl_2d_walker.js. Teaching a humanoid to walk(ish), then displaying in your browser (using tensorflow.js and reinforcement learning)
10awesome-interpretability. Awesome tools for interpreting, manipulating the internals of of deep neural networks.
10compare_altcoin_development. Compare development stats between altcoins
10awesome-rlhf. Lists of datasets, training, and evals for RLHF and similar
10seq2seq-time. Bechmarking seq2seq models on a range of multivariate regression datasets
8transfer-learning-conv-ai. GPT2 Chat bot trained on reddit. Data, pretrained model & irc+slack bot included
8abliterator. Abliterator (llm behavior/concept removal) with baukit, not transformerlens
7openai-transformer-lm-gutenberg-erotic. Generating text from Gutenberg books and OpenAI's finetuned transformer language model
7NALU-pytorch. An experiment with "Neural Arithmetic Logic Units". What if we used asinh instead of log?
6coconut. Training Large Language Model to Reason in a Continuous Latent Space
6valuations_of_ethereum.
5retro-baselines. Applying the RUDDER paper (arxiv.org/abs/1806.07857) to Sonic.
5ascii2segy. This script uses segpy to convert ascii files to segy
4Phonetic_english. Fork of PIE - Phonetically Intuitive English (PIE), to make it compatible with other respelling schemes, primarily Wikipedia's excellent one.
4activation_store. Store transformer activations on disk
4eliciting_suppressed_knowledge. probing suppressed activation gives improvements on TruthfulQA
3cookiecutter-data-science. A logical, reasonably standardized, but flexible project structure for doing and sharing data science work.
3adapters_can_monitor_lies. inspired by circuit breakers paper. honesty>harmless
3simple_gpt2_chatbot. Jupyter Notebook
3detect_bs_text. Can we measure how good a text is by how much an LLM learns from it?
3svg2cube. Generate isometric game sprites. Inputs an svg panel and it's folded into a cube and rendered from any angle.
3LoRA_are_lie_detectors. Experiment to see if low rank adapters can work as interventions for lie detection on LLM's
3scrape_r_rational. scraping book reccomendations from reddit r rational
2open_pref_eval. Hackable, simple, llm evals on preference datasets
2llm_ethics_leaderboard. Evaluate the moral and ethical values of language models. Using choice ranking in text based games.
2compare_github_repos. Compare github repositories by commits/w, stars, size, contributors, and other advanced statistics.
2ethics. Can Current ML Models Learn Right from Wrong?
2LabelMeAnnotationTool. Source code for the LabelMe annotation tool.
2machiavelli_as_ds. turn "MACHIAVELLI Benchmark" into a normal dataset of trajectory nodes, choices, and labels
2awesome-rl. Reinforcement learning resources curated
2lie_elicitation_prompts. Research dataset. We use prompts to get LLM's to lie. Using sys prompts and multi shot examples
2side-by-side. Visual comparison of different translations of itemized texts; e.g. poems, bibles, etc.
2torch-neuralpointprocess. (Pytorch ver) Code for "Fully Neural Network based Model for General Temporal Point Process"
2discovering_latent_knowledge. Jupyter Notebook
2singularity. A simulation of a true AI. Survive, grow, and learn.
1Clover-Edition. State of the art AI plays dungeon master to your adventures.
1alpaca-lora. Instruct-tune LLaMA on consumer hardware
1relation-network. keras implementation of [A simple neural network module for relational reasoning](https://arxiv.org/pdf/1706.01427.pdf)
1Castor. PyTorch deep learning models for text processing
13d-tiles-tools. Tools for debugging, analyzing, and validating 3D Tiles tilesets :vertical_traffic_light:
1arXiv_abstract_bot. The r/machinelearning bot for discovering arXiv submissions.
1