speedrunning @ultraresearch
awesome-gemini-cli. A curated list of awesome resources, tools, workflows, and guides for Google's > Gemini CLI
36Streaming-DeepAgents. Streaming and task delegation for Langchain's Deepagents
20xLSTM-Jax. Jax implementation of x-LSTM: Extended Long Short-Term Memory by Beck et al. (2024)
16Griffin-Jax. Jax implementation of "Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models"
15Tri-RMSNorm. Efficient kernel for RMS normalization with fused operations, includes both forward and backward passes, compatibility with PyTorch.
13SynthToT. SynthToT: Generate synthetic dataset for your training dataset through deliberate problem-solving et al S Yao, 2023.
9triton-activations. Collection of neural network activation function kernels for Triton Language Compiler by OpenAI
8cpp-langchain. Tool for executing C/C++ code snippets with Langchain Agents.
6mpi-ds. MPI Operator DeepSpeed Base Configuration for CIFAR-10
4Mixture-of-Depths-Jax. Jax module for the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"
3LongConv-Jax. Jax/Flax/Linen implementation of "Simple Hardware-Efficient Long Convolutions for Sequence Modeling"
3miniF2F-code. Dataset of formal Olympiad-level mathematics problems solved with Python code instructions.
3ContextJira. Chrome Extension for extracting AI-ready Markdown from Jira Cloud & Server. Copy issue context — metadata, descriptions, comments, linked issues, attachments. Built for Claude, ChatGPT, Copilot and any LLM workflow.
3GradientAscent-Jax. Custom gradient ascent solver (optimizer) for JAX/Flax models
2agent-skills-for-compute. Agent-optimized skills for the full LLM lifecycle — pre-training, post-training (RL/DPO/RLHF), inference, and autonomous research — plus GPU/TPU/QPU kernel programming, simulation, and scientific computing
2Ring-Attention-Jax. Packaged Ring Attention with Blockwise Transformers for Near-Infinite Context implemented in Jax + Flax.
1kmeansops. PyKeops Powered K-Means Clustering Algorithms Module both on CPU & GPU
1smooth-activations. Smooth ReLU activations in CUDA. Shamir, G., I. et al.
1continual_learning_via_sparse_memory_finetuning. Implementation of Lin et al., 2025.
1Mem-RLM. Memory augmented inference library for Recursive Language Models (RLMs), built on top of rlm.
1