This is your work, valued
LlamaGym. Fine-tune LLM agents with online reinforcement learning
1.3kinterrupting-cow. 🐮📢 The first AI voice assistant that interrupts *you*
148sanskrit-ocr. Python
48complexity-scaling. gzip Predicts Data-dependent Scaling Laws
35MindMapResearch. Development of an NLP algorithm to temporally track patient emotional state using digital footprint
13deredact. De-redacting Elon's Email with Character-count Constrained Llama2 Decoding
10gpt2-explorer. gpt-2 latent representation explorer, runs in browser & visualizes similarity as graph connectedness
4SATScale. Towards scalable solutions for boolean satisfiability
3yowza. Automatic generation of listicles from AskReddit threads
3i-am-a-strange-dataset. Repository for "I am a Strange Dataset: Metalinguistic Tests for Language Models"
2ask-a-lecture. Search engine for video lectures that increases information retrieval efficiency by >90%
2celery. Using AI and Big Data to optimize sustainability and profitability for small businesses
2prime-rl. prime-rl is a codebase for decentralized RL training at scale
2basehacks. HTML
2atropos. Atropos is a Language Model Reinforcement Learning Environments framework for collecting and evaluating LLM trajectories through diverse environments
1APCS_Work. All my AP CS work in Java from high school; deleting from local machine
1nano-r1-speedrun. Single File, Single GPU, From Scratch, Efficient, Full Parameter Tuning library for "RL for LLMs"
1BitFix. BitFix is a service that connects open source social good projects that need more hands on deck with volunteers that want to donate their time, skill, and energy to such projects
1tinygrad. You like pytorch? You like micrograd? You love tinygrad! ❤️
1Alleviair. Disaster Alleviation via Air
1vidvas. Sanskrit text database and navigator
1CACR. Cross-modal Attention Congruence Regularization implementation
1FbIrl. Augmented Reality Experiment in Futuristic Social Media
1