Seoul, South Korea

qwopqwop200

Expert
@qwopqwop200

If I have seen further it is by standing on the shoulders of giants

GPTQ-for-LLaMa. 4 bits quantization of LLaMA using GPTQ

3.1k

gptqlora. GPTQLoRA: Efficient Finetuning of Quantized LLMs with GPTQ

101

lama-with-maskdino. automatic image inpainting (lama(with refinement) and maskdino)

47

ko-arena-hard-auto. Ko-Arena-Hard-Auto: An automatic LLM benchmark for Korean

22

stable-diffusion-webui-promptgen-danbooru. stable-diffusion-webui-promptgen

18

GPTQ-for-KoAlpaca. Python

15

SoftPool. softpool implementation(Refining activation downsampling with SoftPool) This is an unofficial implementation. https://arxiv.org/pdf/2101.00440v2.pdf

15

AutoQuant. Python

11

MaxVIT-pytorch. MaxVIT implementation(MaxViT: Multi-Axis Vision Transformer) This is an unofficial implementation. https://arxiv.org/abs/2204.01697

9

llama-danbooru-qlora. Jupyter Notebook

8

ALMA-en2ko. This is repository for ALMA translation models.

8

Neighborhood-Attention-Transformer. NAT implementation(Neighborhood Attention Transformer) This is an unofficial implementation. https://arxiv.org/pdf/2204.07143.pdf

7

pale-transformer-pytorch. Pale Transformer implementation(Pale Transformer: A General Vision Transformer Backbone with Pale-Shaped Attention) This is an unofficial implementation. https://arxiv.org/abs/2112.14000

6

NatIR. NatIR: Image Restoration Using Neighborhood-Attention-Transformer

6

KoLIMA. Jupyter Notebook

4

MLP-Mixer-tf2. MLP-Mixer implementation(MLP-Mixer: An all-MLP Architecture for Vision) This is an unofficial implementation. https://arxiv.org/pdf/2105.01601v1.pdf

3

Magneto-pytorch. Magneto implementation(Foundation Transformers) This is an unofficial implementation. https://arxiv.org/abs/2210.06423

3

Swin-MLP-Mixer. This code is a structure that combines swim-transformer and mlp mixer, and performance may be poor because I didn’t train and test this model

3

Subtitles-generator-with-whisper. Subtitles generator using whisper and translator

3

D-Adaptation-Adan. Adan with D-Adaptation automatic step-sizes

2

transformers-t5. 🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.

2

DynaBOA. Out-of-Domain Human Mesh Reconstruction via Dynamic Bilevel Online Adaptation

1

psnr_ssim_ycbcr. Code for calculating psnr and ssim in y channel in ycbcr.This code is based on BasicSR (https://github.com/xinntao/BasicSR).

1

halonet-tf2. halonet implementation(Scaling Local Self-Attention for Parameter Efficient Visual Backbones) This is an unofficial implementation.https://arxiv.org/pdf/2103.12731v2.pdf

1
24
Apply