This is your work, valued
triton-windows. Fork of the Triton language and compiler for Windows support and easy installation
2kSageAttention. Fork of SageAttention for Windows wheels and easy installation
890transformers-qwen3-moe-fused. Fused Qwen3 MoE layer for faster training, compatible with Transformers, LoRA, bnb 4-bit quant, Unsloth. Also possible to train LoRA over GGUF
259pkuholebackup.
168ComfyUI-RadialAttn. RadialAttention in ComfyUI native workflow
120typeset. 自动修正中文、英文、代码混合排版中的全半角、空格等问题
105pkubbsbackup.
78ACE-Step. Fork of ACE-Step v1.0 for LoRA training with < 10 GB VRAM
70project-asteria. Project Asteria: A Naïve Introductory to Advanced Mathematics and Theoretical Physics for Gaokao Students
44SpargeAttn. Fork of SpargeAttention (SparseSageAttention) for Windows wheels and easy installation
37ComfyUI-FeatherOps. Fast fp16-fp8 mixed precision matmul on RDNA3/3.5 GPUs without native fp8
34read-matrix. Reviews to The Matrix series
21mathstudio-apk-modernized. MathStudio Android APK modernized for Android >= 7
19sageattention-autotune. SageAttention with autotuned block sizes
18evotensile. Evolutionary algorithm for hipBLASLt TensileLite config search
7rdna35-isa-markdown. AMD GPU RDNA3.5 (Strix Halo) instruction set architecture manual in Markdown for AI retrieval
6typeset-rs. 自动修正中文、英文、代码混合排版中的全半角、空格等问题
1