Working on LLMs evaluation 🤖 Aiming to distill human intelligence ðŸ§
CUDA-im2col-conv. CUDA project for uni subject
26GradientAILab. Experiments with Deep Reinforcement Learning
13World-Models. Implementation of World Models: https://worldmodels.github.io.
7humblerl. Straightforward reinforcement learning python framework
7talksh. talksh is a minimal, Unix-friendly command-line tool for running LLM prompts on data. It reads from stdin, calls an LLM, and prints the result to stdout.
2AlphaZero. AlphaZero for Othello, Connect4, GoBang and TicTacToe using Keras and my HumbleRL framework.
1tf_utils. TensorFlow utilities
1