This is your work, valued
rude-carnie. Age detection in Tensorflow
★ 935dliss-tutorial. Tutorial for International Summer School on Deep Learning, 2019
★ 318mint. MinT: Minimal Transformer Library and Tutorials
★ 263textrank-js. TextRank algorithm implementation in Javascript
★ 40sgdtk. A Java library for Stochastic Gradient Descent (SGD)
★ 22arcs-py. Arc-Eager and Arc-Hybrid Greedy Dependency Parser with Dynamic Oracle in Python (with no Dependencies!)
★ 20reserve. FastAPI + WebSockets + SSE service to interface with Triton/Riva ASR
★ 13nbsvm-xl. NBSVM using SGD for fast processing of large datasets
★ 8n3rd. Simple, no frills, neural nets / deep learning in Java
★ 8lit.
★ 5n3rd-cpp. A library for Stochastic Gradient Descent (C++ port)
★ 5lucifR. Lucene Interface for R
★ 4simple-kernel-nashorn. A small, simple kernel example for executing Nashorn JavaScript in Jupyter
★ 2mead-tutorials. Tutorial on using MEAD
★ 2simple-kernel-java. A small, simple kernel example for Jupyter written in Java
★ 2neural-editor. Repository for "Generating Sentences by Editing Prototypes"
★ 1torchure. Various accrued torch utilities including Word2Vec model loader
★ 1Meetups. Presentations for meetups, etc
★ 1grain. Library for reading and processing ML training data.
★ 758dlcalc. random command line tools for deep learning
★ 10ringattention. Large Context Attention
★ 773NeMo-Aligner. Scalable toolkit for efficient model alignment
★ 852direct-preference-optimization. Reference implementation for DPO (Direct Preference Optimization)
★ 2.9kneuronx-distributed. Python
★ 68neuronx-nemo-megatron. Python
★ 39aws-neuron-reference-for-megatron-lm. Python
★ 14optimum-neuron. Training and inference on AWS Trainium and Inferentia chips.
★ 267alpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kUniEval. Repository for EMNLP 2022 Paper: Towards a Unified Multi-Dimensional Evaluator for Text Generation
★ 217Open-Assistant. OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
★ 37kbaize-chatbot. Let ChatGPT teach your own chatbot in hours with a single GPU!
★ 3.2kopenai-cookbook. Examples and guides for using the OpenAI API
★ 75kalexa-teacher-models. Python
★ 363natural-instructions. Expanding natural instructions
★ 1kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30kself-instruct. Aligning pretrained language models with instruction data generated by themselves.
★ 4.6kwikiwiki-dataset. Shell
★ 11deepspeed-sagemaker-example. Jupyter Notebook
★ 22self_dialogue_corpus. The Self-dialogue Corpus - a collection of self-dialogues across music, movies and sports
★ 107chirpycardinal. Stanford's Alexa Prize socialbot
★ 134orconvqa-release. Python
★ 58OAT. An open source toolkit for multimodal generative conversational task assistants, helping assist people with real-world complex tasks
★ 37VANiLLa. A large scale data for answer verbalization for simple natural questions.
★ 9trlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4.8kdialog-inpainting.
★ 97DialogRPT. EMNLP 2020: "Dialogue Response Ranking Training with Large-Scale Human Feedback Data"
★ 345KQAPro_Baselines. Pytorch implementation of baseline models of KQA Pro, a large-scale dataset of complex question answering over knowledge base.
★ 138RL4LMs. A modular RL library to fine-tune language models to human preferences
★ 2.4kacl2022-zerofewshot-tutorial.
★ 290Knover. Large-scale open domain KNOwledge grounded conVERsation system based on PaddlePaddle
★ 669OpenEL_corpus. OpenEL corpus
★ 2fed. Code for SIGdial 2020 paper: Unsupervised Evaluation of Interactive Dialog with DialoGPT (https://arxiv.org/abs/2006.12719)
★ 28SuRE. [EMNLP 2022] Summarization as Indirect Supervision for Relation Extraction (SuRE)
★ 27DST-as-Prompting. Source code for Dialogue State Tracking with a Language Model using Schema-Driven Prompting
★ 66pomdp-baselines. Simple (but often Strong) Baselines for POMDPs in PyTorch, ICML 2022
★ 347architecture-objective. Python
★ 100wikifact. Wikipedia based dataset to train relationship classifiers and fact extraction models
★ 25contrack. The Context Tracking dataset consists of annotated human-human social conversations. Each conversation contains annotations for the people and location entities mentioned, their properties and the relationships between them. The annotated data enables several subtasks like slot tagging, coreference resolution, resolving plural mentions and entity linking.
★ 7Taskmaster. Please see the readme file as well as our 2019 EMNLP paper linked here -->
★ 222EntLM. Codes for "Template-free Prompt Tuning for Few-shot NER".
★ 116massive. Tools and Modeling Code for the MASSIVE dataset
★ 564ReCross. ReCross: Unsupervised Cross-Task Generalization via Retrieval Augmentation
★ 23GenIE. The autoregressive information extraction system GenIE (Generative Information Extraction) implemented in PyTorch.
★ 104Channel-LM-Prompting. An original implementation of "Noisy Channel Language Model Prompting for Few-Shot Text Classification"
★ 130airdialogue. Python
★ 47promptsource. Toolkit for creating, sharing and using natural language prompts.
★ 3kMetaICL. An original implementation of "MetaICL Learning to Learn In Context" by Sewon Min, Mike Lewis, Luke Zettlemoyer and Hannaneh Hajishirzi
★ 274s3prl. Self-Supervised Speech Pre-training and Representation Learning Toolkit
★ 2.6kt-zero. Reproduce results and replicate training fo T0 (Multitask Prompted Training Enables Zero-Shot Task Generalization)
★ 463semsup. Semantic Supervision: Enabling Generalization over Output Spaces
★ 16YaLM-100B. Pretrained language model with 100B parameters
★ 3.8kequinox. Elegant easy-to-use neural networks + scientific computing in JAX. https://docs.kidger.site/equinox/
★ 2.9kOpenPrompt. An Open-Source Framework for Prompt-Learning.
★ 4.9kt-few. Code for T-Few from "Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning"
★ 460FasterTransformer. Transformer related optimization, including BERT, GPT
★ 6.4kmetaseq. Repo for external large-scale work
★ 6.6kConverse. Python
★ 132prompt-tuning. Original Implementation of Prompt Tuning from Lester, et al, 2021
★ 701rebel. REBEL is a seq2seq model that simplifies Relation Extraction (EMNLP 2021).
★ 572CrossWeigh. CrossWeigh: Training Named Entity Tagger from Imperfect Annotations
★ 177contriever. Contriever: Unsupervised Dense Information Retrieval with Contrastive Learning
★ 780annotated-s4. Implementation of https://srush.github.io/annotated-s4
★ 519Few-NERD. Code and data of ACL 2021 paper "Few-NERD: A Few-shot Named Entity Recognition Dataset"
★ 399alexa-end-to-end-slu. This setup allows to train end-to-end neural models for spoken language understanding (SLU).
★ 24pySBD. 🐍💯pySBD (Python Sentence Boundary Disambiguation) is a rule-based sentence boundary detection that works out-of-the-box.
★ 927syntactic-augmentation-nli. Create augmentation examples from MultiNLI by subject-object inversion and passivizing.
★ 17ACE. [ACL-IJCNLP 2021] Automated Concatenation of Embeddings for Structured Prediction
★ 313cista. Cista is a simple, high-performance, zero-copy C++ serialization & reflection library.
★ 2.2kpyannote-audio. Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding
★ 10kvoxpopuli. A large-scale multilingual speech corpus for representation learning, semi-supervised learning and interpretation
★ 574deit. Official DeiT repository
★ 4.4kaudioset_tagging_cnn. Python
★ 1.8kdiy. Do It yourself Dependency Injection in python
★ 20prefix-beam-search. Code for prefix beam search tutorial by @labodk
★ 187extending-jax. Extending JAX with custom C++ and CUDA code
★ 404awesome-multimodal-ml. Reading list for research topics in multimodal machine learning
★ 6.9kDALLE-pytorch. Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
★ 5.6kuncertainty-baselines. High-quality implementations of standard and SOTA methods on a variety of tasks.
★ 1.6kPicking_BERTs_Brain. Code and Data for Picking BERT's Brain: Probing for Linguistic Dependencies in Contextualized Embeddings Using Representational Similarity Analysis by Michael Lepori and R. Thomas McCoy, presented at COLING 2020
★ 4transducer. A Fast Sequence Transducer Implementation with PyTorch Bindings
★ 200EEND. End-to-End Neural Diarization
★ 435jiwer. Evaluate your speech-to-text system with similarity measures such as word error rate (WER)
★ 917WER-in-python. This program calculates the word error rate of hypothesis in ASR and print the aligned result.
★ 158TensorFlowASR. :zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
★ 1kunilm. Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22klibri-light. dataset for lightly supervised training using the librivox audio book recordings. https://librivox.org/.
★ 528iobes. Tool for parsing and converting various span encoding schemes.
★ 23cloud-sql-proxy. A utility for connecting securely to your Cloud SQL instances
★ 1.4kparser. :rocket: State-of-the-art parsers for natural language.
★ 879tetra-tagging. Tetra-Tagging: Word-Synchronous Parsing with Linear-Time Inference
★ 14TreebankPreprocessing. Python scripts preprocessing Penn Treebank and Chinese Treebank
★ 160Better_LSTM_PyTorch. An LSTM in PyTorch with best practices (weight dropout, forget bias, etc.) built-in. Fully compatible with PyTorch LSTM.
★ 133deep_srl. Code and pre-trained model for: Deep Semantic Role Labeling: What Works and What's Next
★ 334bert-stable-fine-tuning. On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines
★ 138self-attentive-parser. High-accuracy NLP parser with models for 11 languages.
★ 911HPSG-Neural-Parser. Source code for "Head-Driven Phrase Structure Grammar Parsing on Penn Treebank" published at ACL 2019
★ 107gap. Gendered Ambiguous Pronouns Shared Task
★ 31rnn_agreement. Evaluating recurrent neural networks on predicting subject-verb agreement dependencies
★ 63blimp. The Benchmark of Linguistic Minimal Pairs
★ 171albert. ALBERT: A Lite BERT for Self-supervised Learning of Language Representations
★ 3.3kpix-plot. A WebGL viewer for UMAP or TSNE-clustered images
★ 649CorefQA. This repo contains the code for ACL2020 paper "Coreference Resolution as Query-based Span Prediction"
★ 139gap-coreference. GAP is a gender-balanced dataset containing 8,908 coreference-labeled pairs of (ambiguous pronoun, antecedent name), sampled from Wikipedia for the evaluation of coreference resolution in practical applications.
★ 228textgrid. A Python module for interacting with Praat TextGrid files. Also includes a class for reading HTK .mlf files into Praat
★ 302e2e-coref. End-to-end Neural Coreference Resolution
★ 525coref-ee. Coreference Resolution With Entity Equalization
★ 40coref. BERT for Coreference Resolution
★ 455ConvLab-2. ConvLab-2: An Open-Source Toolkit for Building, Evaluating, and Diagnosing Dialogue Systems
★ 466code_snippets. Shell
★ 57NLU_datasets_with_task_oriented_dialogue. datasets of natural language understanding and dialogue state tracking
★ 147BlingFire. A lightning fast Finite State machine and REgular expression manipulation library.
★ 1.9kwikiextractor. A tool for extracting plain text from Wikipedia dumps
★ 4kToD-BERT. Pre-Trained Models for ToD-BERT
★ 295elastic. PyTorch elastic training
★ 727text-to-text-transfer-transformer. Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"
★ 6.5kdatasets. 🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
★ 22koos-eval. Repository that accompanies "An Evaluation Dataset for Intent Classification and Out-of-Scope Prediction" (EMNLP 2019)
★ 222conventional-changelog-action. Github Action that generates a changelog with the Conventional Changelog CLI
★ 340self_talk. Code and data for the paper: "Unsupervised Common Sense Question Answering with Self-Talk"
★ 79PyTorch-GAN. PyTorch implementations of Generative Adversarial Networks.
★ 17kStackPropagation-SLU. Open source code for EMNLP-19 Paper "A Stack-Propagation Framework with Token-Level Intent Detection for Spoken Language Understanding".
★ 147hub. A hub for pre-trained models
★ 3end-to-end-SLU. PyTorch code for end-to-end spoken language understanding (SLU) with ASR-based transfer learning
★ 231electra. ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators
★ 2.4kcontextual. How Contextual are Contextualized Word Representations?
★ 43TENER. Codes for "TENER: Adapting Transformer Encoder for Named Entity Recognition"
★ 376nn4nlp-concepts. A repository of concepts related to neural networks for NLP
★ 456trax. Trax — Deep Learning with Clear Code and Speed
★ 8.3kpython_speech_features. This library provides common speech features for ASR including MFCCs and filterbank energies.
★ 2.4kFactorized-TDNN. PyTorch implementation of the Factorized TDNN (TDNN-F) from "Semi-Orthogonal Low-Rank Matrix Factorization for Deep Neural Networks" and Kaldi
★ 149dstc8-schema-guided-dialogue. The Schema-Guided Dialogue Dataset
★ 608ASR. Python
★ 55malmo. Project Malmo is a platform for Artificial Intelligence experimentation and research built on top of Minecraft. We aim to inspire a new generation of research into challenging new problems presented by this unique environment. --- For installation instructions, scroll down to *Getting Started* below, or visit the project page for more information:
★ 4.3kudify. A single model that parses Universal Dependencies across 75 languages. Given a sentence, jointly predicts part-of-speech tags, morphology tags, lemmas, and dependency trees.
★ 225MRDA-Corpus. Utilities for Processing the Meeting Recorder Dialogue Act Corpus
★ 38Switchboard-Corpus. Utilities for Processing the Switchboard Dialogue Act Corpus
★ 73LAMA. LAnguage Model Analysis
★ 1.4kpytorch2keras. PyTorch to Keras model convertor
★ 862addons. Useful extra functionality for TensorFlow 2.x maintained by SIG-addons
★ 1.7kespresso. Espresso: A Fast End-to-End Neural Speech Recognition Toolkit
★ 939tutorial. Tutorial on ‘Graph-Based Meaning Representations: Design and Processing’ (ACL 2019)
★ 113parsimonious. The fastest pure-Python PEG parser I can muster
★ 1.9kbi-tempered-loss. Robust Bi-Tempered Logistic Loss Based on Bregman Divergences. https://arxiv.org/pdf/1906.03361.pdf
★ 147ctrl. Conditional Transformer Language Model for Controllable Generation
★ 1.9kflowseq. Generative Flow based Sequence-to-Sequence Toolkit written in Python.
★ 245DeepLearningExamples. State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
★ 15kML-From-Scratch. Machine Learning From Scratch. Bare bones NumPy implementations of machine learning models and algorithms with a focus on accessibility. Aims to cover everything from linear regression to deep learning.
★ 32kgpu-monitoring-tools. Tools for monitoring NVIDIA GPUs on Linux
★ 1.1khpctl. Sample Mead Configs and train baseline models.
★ 1zeshel. Python
★ 141waliki. A wiki engine powered by Django and Git
★ 309Megatron-LM. Ongoing research training transformer models at scale
★ 17ktrade-dst. Source code for transferable dialogue state generator (TRADE, Wu et al., 2019). https://arxiv.org/abs/1905.08743
★ 393slate. A Super-Lightweight Annotation Tool for Experts: Label text in a terminal with just Python
★ 114pointer_summarizer. pytorch implementation of "Get To The Point: Summarization with Pointer-Generator Networks"
★ 910YouTokenToMe. Unsupervised text tokenizer focused on computational efficiency
★ 980mnist-mc-dropout. model uncertainty using mc dropout
★ 21seq2seq-con. Implementation of "Von Mises-Fisher Loss for Training Sequence to Sequence Models with Continuous Outputs"
★ 77xpctl. track baseline experiments
★ 2transformer_lexical_shortcuts. Codebase accompanying the paper 'Widening the Representation Bottleneck in Neural Machine Translation with Lexical Shortcuts', (Emelin, Denis, Ivan Titov, and Rico Sennrich, Fourth Conference on Machine Translation, Florence, 2019.)
★ 11bertviz. BertViz: Visualize Attention in Transformer Models
★ 8.1kattention-analysis. Jupyter Notebook
★ 474contextual-repr-analysis. A toolkit for evaluating the linguistic knowledge and transferability of contextual representations. Code for "Linguistic Knowledge and Transferability of Contextual Representations" (NAACL 2019).
★ 212MASS. MASS: Masked Sequence to Sequence Pre-training for Language Generation
★ 1.1kXLM. PyTorch original implementation of Cross-lingual Language Model Pretraining.
★ 2.9kxlnet. XLNet: Generalized Autoregressive Pretraining for Language Understanding
★ 6.2kforecast-prometheus. A collection of analysis, and machine learning techniques for time series forecasting w/ Prometheus metrics
★ 157Left2Right-Pointer-Parser. Python
★ 19clarification_question_generation_pytorch. Code and data for the paper: Answer-based Adversarial Training for Generating Clarification Questions
★ 42compare-mt. A tool for holistic analysis of language generations systems
★ 471structural-probes. Codebase for testing whether hidden states of neural networks encode discrete structures.
★ 403Densely-Connected-CNN-with-Multiscale-Feature-Attention. Densely Connected CNN with Multi-scale Feature Attention for Text Classification
★ 133DPCNN. Deep Pyramid Convolutional Neural Networks for Text Categorization in PyTorch
★ 197symphony-mt. Symphony Machine Translation
★ 40dep2label. Dependency Parsing as Sequence Labeling
★ 27ast. Code to train Automatic Speech-to-Text (AST) models
★ 9naacl_transfer_learning_tutorial. Repository of code for the tutorial on Transfer Learning in NLP held at NAACL 2019 in Minneapolis, MN, USA
★ 724nvidia_gpu_prometheus_exporter. NVIDIA GPU Prometheus Exporter
★ 253lm. Python
★ 165pytorch-kaldi. pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label computation, and decoding are performed with the kaldi toolkit.
★ 2.4kmesh. Mesh TensorFlow: Model Parallelism Made Easier
★ 1.6knabu. Code for end-to-end ASR with neural networks, build with TensorFlow
★ 110End-to-end-ASR-Pytorch. This is an open source project (formerly named Listen, Attend and Spell - PyTorch Implementation) for end-to-end ASR implemented with Pytorch, the well known deep learning toolkit.
★ 1.2kListen-Attend-Spell. A PyTorch implementation of Listen, Attend and Spell (LAS), an End-to-End ASR framework.
★ 208Speech-Transformer. A PyTorch implementation of Speech Transformer, an End-to-End ASR with Transformer network on Mandarin Chinese.
★ 810py3nvml. Python 3 Bindings for NVML library. Get NVIDIA GPU status inside your program.
★ 250listen-attend-and-spell. Tensorflow implementation of "Listen, Attend and Spell" authored by William Chan. This project utilizes input pipeline and estimator API of Tensorflow, which makes the training and evaluation truly end-to-end.
★ 90jax. Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
★ 36kgraphpipe-py. GraphPipe for python
★ 41graphpipe-tf-py. GraphPipe helpers for TensorFlow
★ 22model_server. A scalable inference server for models optimized with OpenVINO™
★ 905source-to-image. A tool for building artifacts from source and injecting into container images
★ 2.5kdatasets. TFDS is a collection of datasets ready to use with TensorFlow, Jax, ...
★ 4.6knematus. Open-Source Neural Machine Translation in Tensorflow
★ 805kube-openmpi. Open MPI jobs on Kubernetes
★ 120conversational-datasets. Large datasets for conversational AI
★ 1.4klitbank. Annotated dataset of 100 works of fiction to support tasks in natural language processing and the computational humanities.
★ 381PyTorch-BigGraph. Generate embeddings from large-scale graph-structured data.
★ 3.5k