This is your work, valued

Dogukan Tuna

Expert
@dtunai

speedrunning @ultraresearch

awesome-gemini-cli. A curated list of awesome resources, tools, workflows, and guides for Google's > Gemini CLI

36

Streaming-DeepAgents. Streaming and task delegation for Langchain's Deepagents

20

xLSTM-Jax. Jax implementation of x-LSTM: Extended Long Short-Term Memory by Beck et al. (2024)

16

Griffin-Jax. Jax implementation of "Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models"

15

Tri-RMSNorm. Efficient kernel for RMS normalization with fused operations, includes both forward and backward passes, compatibility with PyTorch.

13

SynthToT. SynthToT: Generate synthetic dataset for your training dataset through deliberate problem-solving et al S Yao, 2023.

9

triton-activations. Collection of neural network activation function kernels for Triton Language Compiler by OpenAI

8

cpp-langchain. Tool for executing C/C++ code snippets with Langchain Agents.

6

mpi-ds. MPI Operator DeepSpeed Base Configuration for CIFAR-10

4

Mixture-of-Depths-Jax. Jax module for the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"

3

LongConv-Jax. Jax/Flax/Linen implementation of "Simple Hardware-Efficient Long Convolutions for Sequence Modeling"

3

miniF2F-code. Dataset of formal Olympiad-level mathematics problems solved with Python code instructions.

3

ContextJira. Chrome Extension for extracting AI-ready Markdown from Jira Cloud & Server. Copy issue context — metadata, descriptions, comments, linked issues, attachments. Built for Claude, ChatGPT, Copilot and any LLM workflow.

3

GradientAscent-Jax. Custom gradient ascent solver (optimizer) for JAX/Flax models

2

agent-skills-for-compute. Agent-optimized skills for the full LLM lifecycle — pre-training, post-training (RL/DPO/RLHF), inference, and autonomous research — plus GPU/TPU/QPU kernel programming, simulation, and scientific computing

2

Ring-Attention-Jax. Packaged Ring Attention with Blockwise Transformers for Near-Infinite Context implemented in Jax + Flax.

1

kmeansops. PyKeops Powered K-Means Clustering Algorithms Module both on CPU & GPU

1

smooth-activations. Smooth ReLU activations in CUDA. Shamir, G., I. et al.

1

continual_learning_via_sparse_memory_finetuning. Implementation of Lin et al., 2025.

1

Mem-RLM. Memory augmented inference library for Recursive Language Models (RLMs), built on top of rlm.

1