Maratyszcza

Elite
@Maratyszcza

PeachPy. x86-64 assembler embedded in Python

2.1k

NNPACK. Acceleration package for neural networks on multi-core CPUs

1.7k

pthreadpool. Portable (POSIX/Windows/Emscripten) thread pool for C/C++

394

FP16. Conversion to/from half-precision floating point formats

384

Opcodes. Database of CPU Opcodes

265

FXdiv. C99/C++ header-only library for division via fixed-point multiplication by inverse

61

psimd. Portable 128-bit SIMD intrinsics

61

caffe-nnpack. Caffe with NNPACK integration

59

FPplus. Scientific library for high-precision computations and research

49

clcc. OpenCL offline compiler

21

confu. Ninja-based configuration system

11

caffe. Caffe: a fast open framework for deep learning.

10

laff-demos. Live demos for Linear Algebra - Foundations to Frontiers course on edX

8

CSE6230. High Performance Computing: Tools and Applications course examples

4

blis-bench. Benchmark of matrix-matrix multiplication implementations for Web browsers

3

vision. Datasets, Transforms and Models specific to Computer Vision

2

cpuinfo. CPU INFOrmation library (x86/x86-64/ARM/ARM64, Linux/Windows/Android/macOS/iOS)

2

onnx. Open Neural Network Exchange

1

stm32_bare_lib. System functions and example code for programming the "Blue Pill" STM32-compatible micro-controller boards.

1

XNNPACK. High-efficiency floating-point neural network inference operators for mobile and Web

1

caffe2. Caffe2 is a lightweight, modular, and scalable deep learning framework.

1

asm.yeppp.info. Run GCC (and other compilers) interactively from your web browser and experiment with its generated code

1

simd. Branch of the spec repo scoped to discussion of SIMD in WebAssembly

1

onnx-tensorrt. ONNX-TensorRT: TensorRT backend for ONNX

1
24
Apply