This is your work, valued

Perth, Australia.

wassname (Michael J Clark)

Elite
@wassname

Machine Learning researcher focused on Pragmatic AI alignment. Based in Perth, Australia. Waiting to be uploaded.

rl-portfolio-management. Attempting to replicate "A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem" https://arxiv.org/abs/1706.10059 (and an openai gym environment)

562

keywordshitter2. A website to find long-tail keywords using search suggestions

221

viz_torch_optim. Videos of deep learning optimizers moving on 3D problem-landscapes

107

attentive-neural-processes. implementing "recurrent attentive neural processes" to forecast power usage (w. LSTM baseline, MCDropout)

100

world-models-sonic-pytorch. Attempt at reinforcement learning with curiosity for Sonic the Hedgehog games. Number 149 on OpenAI retro contest leaderboard, but more work needed

33

satellite_leak_detection. Detect water leaks from satellite images using machine learning

31

thinkpad-x380-yoga-scripts. ThinkPad x380 scripts for GNU/Linux that allow Yoga features (tablet mode, pen mode, auto rotate, auto brightness)

23

pipe-segmentation. Using machine learning to find water pipelines in aerial images with 73% accuracy

22

phoneme2grapheme. Teaching machines to spell with deep learning (acc=>80%) e.g. a model hears "pɹˈaʊd˺ɚ" and writes "prowder" (but it should be "prouder")

19

prob_jsonformer. Generate Structured JSON with probs from Language Models

17

quiet-star. investigate Quiet-STaR paper, and it's thought scratchpad

14

awesome-satellite-imagery-competitions. List of machine learning competitions for satellite imagery and remote sensing.

11

pytorch-pretrained-BERT_horror-genor. PyTorch BERT model - for generating horror

11

rl_2d_walker.js. Teaching a humanoid to walk(ish), then displaying in your browser (using tensorflow.js and reinforcement learning)

10

awesome-interpretability. Awesome tools for interpreting, manipulating the internals of of deep neural networks.

10

compare_altcoin_development. Compare development stats between altcoins

10

awesome-rlhf. Lists of datasets, training, and evals for RLHF and similar

10

seq2seq-time. Bechmarking seq2seq models on a range of multivariate regression datasets

8

transfer-learning-conv-ai. GPT2 Chat bot trained on reddit. Data, pretrained model & irc+slack bot included

8

abliterator. Abliterator (llm behavior/concept removal) with baukit, not transformerlens

7

openai-transformer-lm-gutenberg-erotic. Generating text from Gutenberg books and OpenAI's finetuned transformer language model

7

NALU-pytorch. An experiment with "Neural Arithmetic Logic Units". What if we used asinh instead of log?

6

coconut. Training Large Language Model to Reason in a Continuous Latent Space

6

valuations_of_ethereum.

5

retro-baselines. Applying the RUDDER paper (arxiv.org/abs/1806.07857) to Sonic.

5

ascii2segy. This script uses segpy to convert ascii files to segy

4

Phonetic_english. Fork of PIE - Phonetically Intuitive English (PIE), to make it compatible with other respelling schemes, primarily Wikipedia's excellent one.

4

activation_store. Store transformer activations on disk

4

eliciting_suppressed_knowledge. probing suppressed activation gives improvements on TruthfulQA

3

cookiecutter-data-science. A logical, reasonably standardized, but flexible project structure for doing and sharing data science work.

3

adapters_can_monitor_lies. inspired by circuit breakers paper. honesty>harmless

3

simple_gpt2_chatbot. Jupyter Notebook

3

detect_bs_text. Can we measure how good a text is by how much an LLM learns from it?

3

svg2cube. Generate isometric game sprites. Inputs an svg panel and it's folded into a cube and rendered from any angle.

3

LoRA_are_lie_detectors. Experiment to see if low rank adapters can work as interventions for lie detection on LLM's

3

scrape_r_rational. scraping book reccomendations from reddit r rational

2

open_pref_eval. Hackable, simple, llm evals on preference datasets

2

llm_ethics_leaderboard. Evaluate the moral and ethical values of language models. Using choice ranking in text based games.

2

compare_github_repos. Compare github repositories by commits/w, stars, size, contributors, and other advanced statistics.

2

ethics. Can Current ML Models Learn Right from Wrong?

2

LabelMeAnnotationTool. Source code for the LabelMe annotation tool.

2

machiavelli_as_ds. turn "MACHIAVELLI Benchmark" into a normal dataset of trajectory nodes, choices, and labels

2

awesome-rl. Reinforcement learning resources curated

2

lie_elicitation_prompts. Research dataset. We use prompts to get LLM's to lie. Using sys prompts and multi shot examples

2

side-by-side. Visual comparison of different translations of itemized texts; e.g. poems, bibles, etc.

2

torch-neuralpointprocess. (Pytorch ver) Code for "Fully Neural Network based Model for General Temporal Point Process"

2

discovering_latent_knowledge. Jupyter Notebook

2

singularity. A simulation of a true AI. Survive, grow, and learn.

1

Clover-Edition. State of the art AI plays dungeon master to your adventures.

1

alpaca-lora. Instruct-tune LLaMA on consumer hardware

1

relation-network. keras implementation of [A simple neural network module for relational reasoning](https://arxiv.org/pdf/1706.01427.pdf)

1

Castor. PyTorch deep learning models for text processing

1

3d-tiles-tools. Tools for debugging, analyzing, and validating 3D Tiles tilesets :vertical_traffic_light:

1

arXiv_abstract_bot. The r/machinelearning bot for discovering arXiv submissions.

1