Bay Arena, CA

Lianmin Zheng

Elite
@merrymercy

Engineer

awesome-tensor-compilers. A list of awesome compiler projects and papers for tensor computation and deep learning.

2.8k

tvm-mali. Optimizing Mobile Deep Learning on ARM GPU with TVM

184

goGFS. A simple implementation of the Google File System in golang

161

compiler2017. A compiler for the course Compiler 2017 at ACM Class, SJTU.

81

NALU. Implementation of Neural Arithmetic Logic Units (https://arxiv.org/pdf/1808.00508.pdf)

31

Awesome-Efficient-LLM. A curated list for Efficient Large Language Models

11

tvm. End to end Tensor IR/DSL stack for deploying deep learning workloads to hardwares

10

Awesome-LLM-System-Papers.

8

Wolf-Killer. Werewolf, a board game

5

WSC-and-FVM. A C-minus compiler runs on the CASIO FX-9860 graphing calculator

4

WordKiller. An effective tool to help you to review new words.

3

chain-of-thought-hub. Benchmarking large language models' complex reasoning ability with chain-of-thought prompting

3

Halide. a language for fast, portable data-parallel computation

2

instruct-eval. This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.

2

dtr-prototype. Dynamic Tensor Rematerialization prototype (modified PyTorch) and simulator. Paper: https://arxiv.org/abs/2006.09616

1

Virgo. C++

1

HalideIR. Symbolic Expression and Statement Module for new DSLs

1

dl-system-test. testcase for PPCA 2017 deep learning system

1

FlexFlow. A distributed deep learning framework that supports flexible parallelization strategies.

1

dmlc-core. A common bricks library for building scalable and portable distributed machine learning.

1

tensorflow. An Open Source Machine Learning Framework for Everyone

1

pytorch. Tensors and Dynamic neural networks in Python with strong GPU acceleration

1

ao. PyTorch native quantization and sparsity for training and inference

1
23
Apply