This is your work, valued
embodied ai, autogui...
Shadowrocket-ADBlock-Rules-Forever. 提供多款 Shadowrocket 规则,拥有强劲的广告过滤功能。每日 8 时重新构建规则。
★ 29kVLAExplain. VLA model interpretability tools
★ 177so-novel. 小说下载|网文下载 | 网络小说
★ 7.5klow_cost_robot. Python
★ 3.4kact-plus-plus. Imitation learning algorithms with Co-training for Mobile ALOHA: ACT, Diffusion Policy, VINN
★ 3.6klerobot. 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
★ 26kskyvern. Automate browser based workflows with AI
★ 23kRoboWaiter. 本项目为参加达闼杯“机器人大模型与具身智能挑战赛”的参赛作品。我们的目标是结合前沿的大模型技术和具身智能技术,开发能在模拟的咖啡厅场景中承担服务员角色并自主完成各种具身任务的智能机器人。这里是我们的参赛作品《基于大模型和行为树和生成式具身智能体》的机器人控制端代码。
★ 108LLM-Workshop. LLM Workshop by Sourab Mangrulkar
★ 399FairCLIP. [CVPR 2024] FairCLIP: Harnessing Fairness in Vision-Language Learning
★ 101Whisper-Finetune. Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment
★ 1.2kFunASR. Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
★ 20kCogVLM. a state-of-the-art-level open visual language model | 多模态预训练模型
★ 6.7kWhisper-Finetune. Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment
★ 318UniSpeech. UniSpeech - Large Scale Self-Supervised Learning for Speech
★ 486Qwen-Audio. The official repo of Qwen-Audio (通义千问-Audio) chat & pretrained large audio language model proposed by Alibaba Cloud.
★ 1.9kgenmusic_demo_list. a list of demo websites for automatic music generation research
★ 794LM-exp. LLM experiments done during SERI MATS - focusing on activation steering / interpreting activation spaces
★ 105Humpback. 🐋 An unofficial implementation of Self-Alignment with Instruction Backtranslation.
★ 138LEval. [ACL'24 Outstanding] Data and code for L-Eval, a comprehensive long context language models evaluation benchmark
★ 406AutoAgents. [IJCAI 2024] Generate different roles for GPTs to form a collaborative entity for complex tasks.
★ 1.5kToolQA. ToolQA, a new dataset to evaluate the capabilities of LLMs in answering challenging questions with external tools. It offers two levels (easy/hard) across eight real-life scenarios.
★ 286ToolAlpaca. the official code for "ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases"
★ 880serf. Service orchestration and management tool.
★ 6.1kconsul. Consul is a distributed, highly available, and data center aware solution to connect and configure applications across dynamic, distributed infrastructure.
★ 30kknowledge_distillation. Repository for "Propagating Knowledge Updates to LMs Through Distillation" (NeurIPS 2023).
★ 27OpenDelta. A plug-and-play library for parameter-efficient-tuning (Delta Tuning)
★ 1kLOMO. LOMO: LOw-Memory Optimization
★ 993LLM-ToolMaker. Jupyter Notebook
★ 1.1kMNBVC. MNBVC(Massive Never-ending BT Vast Chinese corpus)超大规模中文语料集。对标chatGPT训练的40T数据。MNBVC数据集不但包括主流文化,也包括各个小众文化甚至火星文的数据。MNBVC数据集包括新闻、作文、小说、书籍、杂志、论文、台词、帖子、wiki、古诗、歌词、商品介绍、笑话、糗事、聊天记录等一切形式的纯文本中文数据。
★ 4.2kLLM-QAT. Code repo for the paper "LLM-QAT Data-Free Quantization Aware Training for Large Language Models"
★ 328lmdeploy. LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
★ 8kALCE. [EMNLP 2023] Enabling Large Language Models to Generate Text with Citations. Paper: https://arxiv.org/abs/2305.14627
★ 523FlagEmbedding. Retrieval and Retrieval-augmented LLMs
★ 12kOpenFace. OpenFace – a state-of-the art tool intended for facial landmark detection, head pose estimation, facial action unit recognition, and eye-gaze estimation.
★ 7.7kllm-action. 本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)
★ 25kFirefly. Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型
★ 6.6kAwesome-Chinese-LLM. 整理开源的中文大语言模型,以规模较小、可私有化部署、训练成本较低的模型为主,包括底座模型,垂直领域微调及应用,数据集与教程等。
★ 23kaliendao. huggingface mirror download
★ 588DISC-MedLLM. Repository of DISC-MedLLM, it is a comprehensive solution that leverages Large Language Models (LLMs) to provide accurate and truthful medical response in end-to-end conversational healthcare services.
★ 565vllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kPromptCraft-Robotics. Community for applying LLMs to robotics and a robot simulator with ChatGPT integration
★ 2.1kEverything-LLMs-And-Robotics. The world's largest GitHub Repository for LLMs + Robotics
★ 849Awesome-LLM-Robotics. A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites
★ 4.4kROS-LLM. ROS-LLM is a framework designed for embodied intelligence applications in ROS. It allows natural language interactions and leverages Large Language Models (LLMs) for decision-making and robot control. With an easy configuration process, this framework allows for swift integration, enabling your robot to operate with it in as little as ten minutes.
★ 819AISystem. AISystem 主要是指AI系统,包括AI芯片、AI编译器、AI推理和训练框架等AI全栈底层技术
★ 17kVIMA. Official Algorithm Implementation of ICML'23 Paper "VIMA: General Robot Manipulation with Multimodal Prompts"
★ 855text-generation-inference. Large Language Model Text Generation Inference
★ 11kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kLLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 25kCPM-Live. Live Training for Open-source Big Models
★ 499DeepSpeedExamples. Example models using DeepSpeed
★ 6.8khelm. Holistic Evaluation of Language Models (HELM) is an open source Python framework created by the Center for Research on Foundation Models (CRFM) at Stanford for holistic, reproducible and transparent evaluation of foundation models, including large language models (LLMs) and multimodal models.
★ 2.9kChatGLM-6B. ChatGLM-6B: An Open Bilingual Dialogue Language Model | 开源双语对话语言模型
★ 41kchatgpt-tool-hub. An open-source chatgpt tool ecosystem where you can combine tools with chatgpt and use natural language to do anything.
★ 1.3kOpenChineseLLaMA. Chinese large language model base generated through incremental pre-training on Chinese datasets
★ 239MOSS. An open-source tool-augmented conversational language model from Fudan University
★ 12kBMTools. Tool Learning for Big Models, Open-Source Solutions of ChatGPT-Plugins
★ 2.8kbabyagi. Python
★ 22kFlagInstruct.
★ 173Alpaca-CoT. We unified the interfaces of instruction-tuning data (e.g., CoT data), multiple LLMs and parameter-efficient methods (e.g., lora, p-tuning) together for easy use. We welcome open-source enthusiasts to initiate any meaningful PR on this repo and integrate as many LLM related technologies as possible. 我们打造了方便研究人员上手和使用大模型等微调平台,我们欢迎开源爱好者发起任何有意义的pr!
★ 2.8kLlama-X. Open Academic Research on Improving LLaMA to SOTA LLM
★ 1.6kzero_nlp. 中文nlp解决方案(大模型、数据、模型、训练、推理)
★ 3.8knanoGPT. The simplest, fastest repository for training/finetuning medium-sized GPTs.
★ 62kChatPaper. Use ChatGPT to summarize the arXiv papers. 全流程加速科研,利用chatgpt进行论文全文总结+专业翻译+润色+审稿+审稿回复
★ 20ksimpletransformers. Transformers for Information Retrieval, Text Classification, NER, QA, Language Modelling, Language Generation, T5, Multi-Modal, and Conversational AI
★ 4.3kclassifier-multi-label. 多标签文本分类,多标签分类,文本分类, multi-label, classifier, text classification, BERT, seq2seq,attention, multi-label-classification
★ 806funNLP. 中英文敏感词、语言检测、中外手机/电话归属地/运营商查询、名字推断性别、手机号抽取、身份证抽取、邮箱抽取、中日文人名库、中文缩写库、拆字词典、词汇情感值、停用词、反动词表、暴恐词表、繁简体转换、英文模拟中文发音、汪峰歌词生成器、职业名称词库、同义词库、反义词库、否定词库、汽车品牌词库、汽车零件词库、连续英文切割、各种中文词向量、公司名字大全、古诗词库、IT词库、财经词库、成语词库、地名词库、历史名人词库、诗词词库、医学词库、饮食词库、法律词库、汽车词库、动物词库、中文聊天语料、中文谣言数据、百度中文问答数据集、句子相似度匹配算法集合、bert资源、文本生成&摘要相关工具、cocoNLP信息抽取工具、国内电话号码正则匹配、清华大学XLORE:中英文跨语言百科知识图谱、清华大学人工智能技术系列报告、自然语言生成、NLU太难了系列、自动对联数据及机器人、用户名黑名单列表、罪名法务名词及分类模型、微信公众号语料、cs224n深度学习自然语言处理课程、中文手写汉字识别、中文自然语言处理 语料/数据集、变量命名神器、分词语料库+代码、任务型对话英文数据集、ASR 语音数据集 + 基于深度学习的中文语音识别系统、笑声检测器、Microsoft多语言数字/单位/如日期时间识别包、中华新华字典数据库及api(包括常用歇后语、成语、词语和汉字)、文档图谱自动生成、SpaCy 中文模型、Common Voice语音识别数据集新版、神经网络关系抽取、基于bert的命名实体识别、关键词(Keyphrase)抽取包pke、基于医疗领域知识图谱的问答系统、基于依存句法与语义角色标注的事件三元组抽取、依存句法分析4万句高质量标注数据、cnocr:用来做中文OCR的Python3包、中文人物关系知识图谱项目、中文nlp竞赛项目及代码汇总、中文字符数据、speech-aligner: 从“人声语音”及其“语言文本”产生音素级别时间对齐标注的工具、AmpliGraph: 知识图谱表示学习(Python)库:知识图谱概念链接预测、Scattertext 文本可视化(python)、语言/知识表示工具:BERT & ERNIE、中文对比英文自然语言处理NLP的区别综述、Synonyms中文近义词工具包、HarvestText领域自适应文本挖掘工具(新词发现-情感分析-实体链接等)、word2word:(Python)方便易用的多语言词-词对集:62种语言/3,564个多语言对、语音识别语料生成工具:从具有音频/字幕的在线视频创建自动语音识别(ASR)语料库、构建医疗实体识别的模型(包含词典和语料标注)、单文档非监督的关键词抽取、Kashgari中使用gpt-2语言模型、开源的金融投资数据提取工具、文本自动摘要库TextTeaser: 仅支持英文、人民日报语料处理工具集、一些关于自然语言的基本模型、基于14W歌曲知识库的问答尝试--功能包括歌词接龙and已知歌词找歌曲以及歌曲歌手歌词三角关系的问答、基于Siamese bilstm模型的相似句子判定模型并提供训练数据集和测试数据集、用Transformer编解码模型实现的根据Hacker News文章标题自动生成评论、用BERT进行序列标记和文本分类的模板代码、LitBank:NLP数据集——支持自然语言处理和计算人文学科任务的100部带标记英文小说语料、百度开源的基准信息抽取系统、虚假新闻数据集、Facebook: LAMA语言模型分析,提供Transformer-XL/BERT/ELMo/GPT预训练语言模型的统一访问接口、CommonsenseQA:面向常识的英文QA挑战、中文知识图谱资料、数据及工具、各大公司内部里大牛分享的技术文档 PDF 或者 PPT、自然语言生成SQL语句(英文)、中文NLP数据增强(EDA)工具、英文NLP数据增强工具 、基于医药知识图谱的智能问答系统、京东商品知识图谱、基于mongodb存储的军事领域知识图谱问答项目、基于远监督的中文关系抽取、语音情感分析、中文ULMFiT-情感分析-文本分类-语料及模型、一个拍照做题程序、世界各国大规模人名库、一个利用有趣中文语料库 qingyun 训练出来的中文聊天机器人、中文聊天机器人seqGAN、省市区镇行政区划数据带拼音标注、教育行业新闻语料库包含自动文摘功能、开放了对话机器人-知识图谱-语义理解-自然语言处理工具及数据、中文知识图谱:基于百度百科中文页面-抽取三元组信息-构建中文知识图谱、masr: 中文语音识别-提供预训练模型-高识别率、Python音频数据增广库、中文全词覆盖BERT及两份阅读理解数据、ConvLab:开源多域端到端对话系统平台、中文自然语言处理数据集、基于最新版本rasa搭建的对话系统、基于TensorFlow和BERT的管道式实体及关系抽取、一个小型的证券知识图谱/知识库、复盘所有NLP比赛的TOP方案、OpenCLaP:多领域开源中文预训练语言模型仓库、UER:基于不同语料+编码器+目标任务的中文预训练模型仓库、中文自然语言处理向量合集、基于金融-司法领域(兼有闲聊性质)的聊天机器人、g2pC:基于上下文的汉语读音自动标记模块、Zincbase 知识图谱构建工具包、诗歌质量评价/细粒度情感诗歌语料库、快速转化「中文数字」和「阿拉伯数字」、百度知道问答语料库、基于知识图谱的问答系统、jieba_fast 加速版的jieba、正则表达式教程、中文阅读理解数据集、基于BERT等最新语言模型的抽取式摘要提取、Python利用深度学习进行文本摘要的综合指南、知识图谱深度学习相关资料整理、维基大规模平行文本语料、StanfordNLP 0.2.0:纯Python版自然语言处理包、NeuralNLP-NeuralClassifier:腾讯开源深度学习文本分类工具、端到端的封闭域对话系统、中文命名实体识别:NeuroNER vs. BertNER、新闻事件线索抽取、2019年百度的三元组抽取比赛:“科学空间队”源码、基于依存句法的开放域文本知识三元组抽取和知识库构建、中文的GPT2训练代码、ML-NLP - 机器学习(Machine Learning)NLP面试中常考到的知识点和代码实现、nlp4han:中文自然语言处理工具集(断句/分词/词性标注/组块/句法分析/语义分析/NER/N元语法/HMM/代词消解/情感分析/拼写检查、XLM:Facebook的跨语言预训练语言模型、用基于BERT的微调和特征提取方法来进行知识图谱百度百科人物词条属性抽取、中文自然语言处理相关的开放任务-数据集-当前最佳结果、CoupletAI - 基于CNN+Bi-LSTM+Attention 的自动对对联系统、抽象知识图谱、MiningZhiDaoQACorpus - 580万百度知道问答数据挖掘项目、brat rapid annotation tool: 序列标注工具、大规模中文知识图谱数据:1.4亿实体、数据增强在机器翻译及其他nlp任务中的应用及效果、allennlp阅读理解:支持多种数据和模型、PDF表格数据提取工具 、 Graphbrain:AI开源软件库和科研工具,目的是促进自动意义提取和文本理解以及知识的探索和推断、简历自动筛选系统、基于命名实体识别的简历自动摘要、中文语言理解测评基准,包括代表性的数据集&基准模型&语料库&排行榜、树洞 OCR 文字识别 、从包含表格的扫描图片中识别表格和文字、语声迁移、Python口语自然语言处理工具集(英文)、 similarity:相似度计算工具包,java编写、海量中文预训练ALBERT模型 、Transformers 2.0 、基于大规模音频数据集Audioset的音频增强 、Poplar:网页版自然语言标注工具、图片文字去除,可用于漫画翻译 、186种语言的数字叫法库、Amazon发布基于知识的人-人开放领域对话数据集 、中文文本纠错模块代码、繁简体转换 、 Python实现的多种文本可读性评价指标、类似于人名/地名/组织机构名的命名体识别数据集 、东南大学《知识图谱》研究生课程(资料)、. 英文拼写检查库 、 wwsearch是企业微信后台自研的全文检索引擎、CHAMELEON:深度学习新闻推荐系统元架构 、 8篇论文梳理BERT相关模型进展与反思、DocSearch:免费文档搜索引擎、 LIDA:轻量交互式对话标注工具 、aili - the fastest in-memory index in the East 东半球最快并发索引 、知识图谱车音工作项目、自然语言生成资源大全 、中日韩分词库mecab的Python接口库、中文文本摘要/关键词提取、汉字字符特征提取器 (featurizer),提取汉字的特征(发音特征、字形特征)用做深度学习的特征、中文生成任务基准测评 、中文缩写数据集、中文任务基准测评 - 代表性的数据集-基准(预训练)模型-语料库-baseline-工具包-排行榜、PySS3:面向可解释AI的SS3文本分类器机器可视化工具 、中文NLP数据集列表、COPE - 格律诗编辑程序、doccano:基于网页的开源协同多语言文本标注工具 、PreNLP:自然语言预处理库、简单的简历解析器,用来从简历中提取关键信息、用于中文闲聊的GPT2模型:GPT2-chitchat、基于检索聊天机器人多轮响应选择相关资源列表(Leaderboards、Datasets、Papers)、(Colab)抽象文本摘要实现集锦(教程 、词语拼音数据、高效模糊搜索工具、NLP数据增广资源集、微软对话机器人框架 、 GitHub Typo Corpus:大规模GitHub多语言拼写错误/语法错误数据集、TextCluster:短文本聚类预处理模块 Short text cluster、面向语音识别的中文文本规范化、BLINK:最先进的实体链接库、BertPunc:基于BERT的最先进标点修复模型、Tokenizer:快速、可定制的文本词条化库、中文语言理解测评基准,包括代表性的数据集、基准(预训练)模型、语料库、排行榜、spaCy 医学文本挖掘与信息提取 、 NLP任务示例项目代码集、 python拼写检查库、chatbot-list - 行业内关于智能客服、聊天机器人的应用和架构、算法分享和介绍、语音质量评价指标(MOSNet, BSSEval, STOI, PESQ, SRMR)、 用138GB语料训练的法文RoBERTa预训练语言模型 、BERT-NER-Pytorch:三种不同模式的BERT中文NER实验、无道词典 - 有道词典的命令行版本,支持英汉互查和在线查询、2019年NLP亮点回顾、 Chinese medical dialogue data 中文医疗对话数据集 、最好的汉字数字(中文数字)-阿拉伯数字转换工具、 基于百科知识库的中文词语多词义/义项获取与特定句子词语语义消歧、awesome-nlp-sentiment-analysis - 情感分析、情绪原因识别、评价对象和评价词抽取、LineFlow:面向所有深度学习框架的NLP数据高效加载器、中文医学NLP公开资源整理 、MedQuAD:(英文)医学问答数据集、将自然语言数字串解析转换为整数和浮点数、Transfer Learning in Natural Language Processing (NLP) 、面向语音识别的中文/英文发音辞典、Tokenizers:注重性能与多功能性的最先进分词器、CLUENER 细粒度命名实体识别 Fine Grained Named Entity Recognition、 基于BERT的中文命名实体识别、中文谣言数据库、NLP数据集/基准任务大列表、nlp相关的一些论文及代码, 包括主题模型、词向量(Word Embedding)、命名实体识别(NER)、文本分类(Text Classificatin)、文本生成(Text Generation)、文本相似性(Text Similarity)计算等,涉及到各种与nlp相关的算法,基于keras和tensorflow 、Python文本挖掘/NLP实战示例、 Blackstone:面向非结构化法律文本的spaCy pipeline和NLP模型通过同义词替换实现文本“变脸” 、中文 预训练 ELECTREA 模型: 基于对抗学习 pretrain Chinese Model 、albert-chinese-ner - 用预训练语言模型ALBERT做中文NER 、基于GPT2的特定主题文本生成/文本增广、开源预训练语言模型合集、多语言句向量包、编码、标记和实现:一种可控高效的文本生成方法、 英文脏话大列表 、attnvis:GPT2、BERT等transformer语言模型注意力交互可视化、CoVoST:Facebook发布的多语种语音-文本翻译语料库,包括11种语言(法语、德语、荷兰语、俄语、西班牙语、意大利语、土耳其语、波斯语、瑞典语、蒙古语和中文)的语音、文字转录及英文译文、Jiagu自然语言处理工具 - 以BiLSTM等模型为基础,提供知识图谱关系抽取 中文分词 词性标注 命名实体识别 情感分析 新词发现 关键词 文本摘要 文本聚类等功能、用unet实现对文档表格的自动检测,表格重建、NLP事件提取文献资源列表 、 金融领域自然语言处理研究资源大列表、CLUEDatasetSearch - 中英文NLP数据集:搜索所有中文NLP数据集,附常用英文NLP数据集 、medical_NER - 中文医学知识图谱命名实体识别 、(哈佛)讲因果推理的免费书、知识图谱相关学习资料/数据集/工具资源大列表、Forte:灵活强大的自然语言处理pipeline工具集 、Python字符串相似性算法库、PyLaia:面向手写文档分析的深度学习工具包、TextFooler:针对文本分类/推理的对抗文本生成模块、Haystack:灵活、强大的可扩展问答(QA)框架、中文关键短语抽取工具
★ 82kGLM-130B. GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
★ 7.7knlp-gym. NLPGym - A toolkit to develop RL agents to solve NLP tasks.
★ 203nlpcda. 一键中文数据增强包 ; NLP数据增强、bert数据增强、EDA:pip install nlpcda
★ 1.9kData-Technology-Books. "One person's data is another person's noise." ― K.C. Cole
★ 157modAL. A modular active learning framework for Python
★ 2.4kALFramework. 主动学习(Active Learning)框架,实现了多个主动学习策略,包括:熵(Entropy)、最大梯度改变(ECG)等。
★ 86labelit. labelit, label tool with active learning, for classification task. 自动标注,基于主动学习,边标注边学习,减少人工标注量。
★ 31Similarity. 文本相似度算法
★ 40awesome-AI-cheatsheets. AI与数据科学各类工具库速查表与参考代码
★ 683machine-learning-yearning-cn. Machine Learning Yearning 中文版 - 《机器学习训练秘籍》 - Andrew Ng 著
★ 7.8kGPLinker_torch. CMeIE/CBLUE/CHIP/实体关系抽取/SPO抽取
★ 243Fengshenbang-LM. Fengshenbang-LM(封神榜大模型)是IDEA研究院认知计算与自然语言研究中心主导的大模型开源体系,成为中文AIGC和认知智能的基础设施。
★ 4.1kmrc-for-flat-nested-ner. Code for ACL 2020 paper `A Unified MRC Framework for Named Entity Recognition`
★ 678nlp-tutorial. Natural Language Processing Tutorial for Deep Learning Researchers
★ 1.1kroberta_zh. RoBERTa中文预训练模型: RoBERTa for Chinese
★ 2.8kalbert_pytorch. A Lite Bert For Self-Supervised Learning Language Representations
★ 714R-BERT. Pytorch implementation of R-BERT: "Enriching Pre-trained Language Model with Entity Information for Relation Classification"
★ 362fastmoe. A fast MoE impl for PyTorch
★ 1.9kDeepIE. DeepIE: Deep Learning for Information Extraction
★ 1.9kDeepNER. 天池中药说明书实体识别挑战冠军方案;中文命名实体识别;NER; BERT-CRF & BERT-SPAN & BERT-MRC;Pytorch
★ 967CLUENER2020. CLUENER2020 中文细粒度命名实体识别 Fine Grained Named Entity Recognition
★ 1.5kfedlearner. A multi-party collaborative machine learning framework
★ 902kg-bert. KG-BERT: BERT for Knowledge Graph Completion
★ 796nebula. A distributed, fast open-source graph database featuring horizontal scalability and high availability
★ 12kCogDL. CogDL: A Comprehensive Library for Graph Deep Learning (WWW 2023)
★ 1.8kbert-utils. 一行代码使用BERT生成句向量,BERT做文本分类、文本相似度计算
★ 1.7knlp_chinese_corpus. 大规模中文自然语言处理语料 Large Scale Chinese Corpus for NLP
★ 9.9kmedical-books. Open sourece medical books in LaTeX. LaTeX写的中文开源医学书籍
★ 717mt-dnn. Multi-Task Deep Neural Networks for Natural Language Understanding
★ 2.3kplato. 腾讯高性能分布式图计算框架Plato
★ 1.9kalbert_zh. A LITE BERT FOR SELF-SUPERVISED LEARNING OF LANGUAGE REPRESENTATIONS, 海量中文预训练ALBERT模型
★ 4ktypedb. TypeDB: Built for systems, not records
★ 4.4kChinese-XLNet. Pre-Trained Chinese XLNet(中文XLNet预训练模型)
★ 1.6kMedicalNet. Many studies have shown that the performance on deep learning is significantly affected by volume of training data. The MedicalNet project provides a series of 3D-ResNet pre-trained models and relative code.
★ 2.2kgraphvite. GraphVite: A General and High-performance Graph Embedding System
★ 1.3kNLP-progress. Repository to track the progress in Natural Language Processing (NLP), including the datasets and the current state-of-the-art for the most common NLP tasks.
★ 23kdrools-demos. Drools 规则引擎示例
★ 12drools. rules engine
★ 306text_classification. all kinds of text classification models and more with deep learning
★ 7.9k996.ICU. Repo for counting stars and contributing. Press F to pay respect to glorious developers.
★ 277kPyTorch-BigGraph. Generate embeddings from large-scale graph-structured data.
★ 3.5kIcarus. 🕊️ An opensource community/forum project write with python3 aiohttp and vue.js. 一个开源的社区程序
★ 643Doctor. 基于知识图谱的医学诊断系统。Medical Diagnosis System Based on Knowledge Map.(欢迎Star,🚫禁止Fork)
★ 347bulma-product. website template using bulma css adapted from https://dansup.github.io/bulma-templates/
★ 1ChineseTextualInference. ChineseTextualInference project including chinese corpus build and inferecence model, 中文文本推断项目,包括88万文本蕴含中文文本蕴含数据集的翻译与构建,基于深度学习的文本蕴含判定模型构建.
★ 174vuescroll. A customizable scrollbar plugin based on vue.js for PC , mobile phone, touch screen, laptop.
★ 1.3kdrf-Vue-website. 编程案例分享网站(基于Django-REST-Framework和Vue.js)
★ 128deep-learning-drizzle. Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
★ 13kOpenHowNet. Core Data of HowNet and OpenHowNet Python API
★ 639MachineLearningNote. 用python实现机器学习各种经典算法
★ 40bert. TensorFlow code and pre-trained models for BERT
★ 40kgraph_nets. Build Graph Nets in Tensorflow
★ 5.4kCEC-Corpus. :books:中文突发事件语料库(Chinese Emergency Corpus)-上海大学-语义智能实验室
★ 721NER. 基于tensorflow深度学习的中文的命名实体识别
★ 1.1kOpenNRE. An Open-Source Package for Neural Relation Extraction (NRE)
★ 4.5kawesome-nlp. :book: A curated list of resources dedicated to Natural Language Processing (NLP)
★ 19kchinese-word2vec. word2vec/glove/swivel binary file on chinese corpus
★ 403ChineseNER. A neural network model for Chinese named entity recognition
★ 1.8kInformation-Extraction-Chinese. Chinese Named Entity Recognition with IDCNN/biLSTM+CRF, and Relation Extraction with biGRU+2ATT 中文实体识别与关系提取
★ 2.3knre. cnn lstm的关系识别 ,keras实现 网络结构简单.多分类
★ 20RE. 关系抽取实验
★ 32LTP_Python_Interface. 根据自己搭的 LTP 服务器,实现:分词、词性标注、命名实体识别、依存句法分析、语义角色标、命名实体的抽取:人名,地名,机构名、三元组的抽取:主谓宾,动宾关系,介宾关系,(实体1,关系,实体2)
★ 145CausalityEventExtraction. Causality event extraction demo project including casual patterns and experiment on large scale corpus. 基于因果关系知识库的因果事件图谱实验项目,本项目罗列了因果显式表达的几种模式,基于这种模式和大规模语料,再经过融合等操作,可形成因果事件图谱。
★ 433open-entity-relation-extraction. Knowledge triples extraction and knowledge base construction based on dependency syntax for open domain text.
★ 537zhopenie. Chinese Open Information Extraction (Tree-based Triple Relation Extraction Module)
★ 116Competition_CAIL. 2018中国‘法研杯’法律智能挑战赛(CAIL2018)个人作品
★ 225chatbot_by_similarity. 根据文本相似度实现问答的聊天机器人(简单版)
★ 51DS_CTT. Distant supervision for Chinese Temporal Tagging
★ 9old_rex. REx: Relation Extraction. Modernized re-write of the code in the master's thesis: "Relation Extraction using Distant Supervision, SVMs, and Probabalistic First-Order Logic"
★ 22factoid_QA_with_distant_spervision. Codes for "Factoid Question Answering With Distant Supervision"
★ 7DeepKE. [EMNLP 2022] An Open Toolkit for Knowledge Graph Extraction and Construction
★ 4.5kTransX. Trans系列之TransE and TransH, PTransE
★ 31spellchecker. 拼写纠错-基于lucene-ngram实现拼写纠错
★ 9KRLPapers. Must-read papers on knowledge representation learning (KRL) / knowledge embedding (KE)
★ 1.5kawesome-graph-classification. A collection of important graph embedding, classification and representation learning papers with implementations.
★ 4.8kAgriculture_KnowledgeGraph. 农业知识图谱(AgriKG):农业领域的信息检索,命名实体识别,关系抽取,智能问答,辅助决策
★ 4.4kfact_triple_extraction. 使用句法依存分析抽取事实三元组
★ 331bulma-vuejs-demo-website. A demo website based on framework Bulma (css) & vuejs (JS)
★ 49dl_class_2017_May. 七月算法深度学习五月班课件
★ 19TransE. A TensorFlow implementation of TransE model
★ 200architect-awesome. 后端架构师技术图谱
★ 61kNRLPapers. Must-read papers on network representation learning (NRL)/network embedding (NE)
★ 12gcn. Implementation of Graph Convolutional Networks in TensorFlow
★ 7.4kawesome-embedding-models. A curated list of awesome embedding models tutorials, projects and communities.
★ 1.8kKB2E. Knowledge Graph Embeddings including TransE, TransH, TransR and PTransE
★ 1.4kawesome-network-embedding. A curated list of network embedding techniques.
★ 2.6kneo4j-timetree. Java and REST APIs for working with time-representing tree in Neo4j
★ 210neo4j-nlp. NLP Capabilities in Neo4j
★ 345zh-NER-TF. A very simple BiLSTM-CRF model for Chinese Named Entity Recognition 中文命名实体识别 (TensorFlow)
★ 2.3kSynonyms. :herb: 中文近义词:聊天机器人,智能问答工具包
★ 5.1kAwesome-Chinese-NLP. A curated list of resources for Chinese NLP 中文自然语言处理相关资料
★ 7.9kTensorflow-Tutorial. Tensorflow tutorial from basic to hard, 莫烦Python 中文AI教学
★ 4.3kawesome-rnn. Recurrent Neural Network - A curated list of resources dedicated to RNN
★ 6.2klectures. Oxford Deep NLP 2017 course
★ 16kapi_ner. API for Tensorflow model in Flask
★ 103RNNSharp. RNNSharp is a toolkit of deep recurrent neural network which is widely used for many different kinds of tasks, such as sequence labeling, sequence-to-sequence and so on. It's written by C# language and based on .NET framework 4.6 or above versions. RNNSharp supports many different types of networks, such as forward and bi-directional network, sequence-to-sequence network, and different types of layers, such as LSTM, Softmax, sampled Softmax and others.
★ 288kcws. Deep Learning Chinese Word Segment
★ 2.1kTensorflow-Tutorial. Some interesting TensorFlow tutorials for beginners.
★ 894d3-dependency-parse-tree. Dependency parse tree visualization using D3.js
★ 7d3-cloud. Create word clouds in JavaScript.
★ 3.9kwechat-admin. Wechat Management System
★ 1.7kbulma-templates. free flexbox templates built with the bulma css framework
★ 3.3kbulmaswatch. Themes for Bulma
★ 1.6kbuefy. Lightweight UI components for Vue.js based on Bulma
★ 9.5kelectron-vue-admin. vue electron admin template web: http://panjiachen.github.io/vue-admin-template
★ 3.2kadminify. An Admin Dashboard based on Vuetify material
★ 958iview-admin. Vue 2.0 admin management system template based on iView
★ 16kdejavu. A Web UI for Elasticsearch and OpenSearch: Import, browse and edit data with rich filters and query views, create reference search UIs.
★ 8.5kQA-Snake. 基于多搜索引擎和深度学习技术的自动问答
★ 644imbalanced-learn. A Python Package to Tackle the Curse of Imbalanced Datasets in Machine Learning
★ 7.1kInterview. Interview = 简历指南 + 算法题 + 八股文 + 源码分析
★ 9kailearning. AiLearning:数据分析+机器学习实战+线性代数+PyTorch+NLTK+TF2
★ 42kneo4jd3. Neo4j graph visualization using D3.js
★ 1.5klavas. 基于 Vue 的 PWA 解决方案,帮助开发者快速搭建 PWA 应用,解决接入 PWA 的各种问题
★ 1.9kvue-element-admin. :tada: A magical vue admin https://panjiachen.github.io/vue-element-admin
★ 90kvue-admin. We are refactoring it, using the latest Vue and Bulma. WIP
★ 9.3kvue2-elm. Large single page application with 45 pages built on vue2 + vuex. 基于 vue2 + vuex 构建一个具有 45 个页面的大型单页面应用
★ 41kcayley. An open-source graph database
★ 15kFileSaver.js. An HTML5 saveAs() FileSaver implementation
★ 22kkg-beijing. 北京知识图谱学习小组
★ 1.7khao. 好东西传送门
★ 1.4kjanusgraph. JanusGraph: an open-source, distributed graph database
★ 5.8kAR.js. Efficient Augmented Reality for the Web - 60fps on mobile!
★ 16kakka-http-quickstart-scala.g8. Scala
★ 125neo4j-graphql-cli. Deploy a Neo4j backed GraphQL API based on your GraphQL schema
★ 61JavaEETest. Spring、SpringMVC、MyBatis、Spring Boot案例
★ 3.5kneural-networks-and-deep-learning. Code samples for my book "Neural Networks and Deep Learning"
★ 18kfreecodecamp.cn. FCC China open source codebase and curriculum. Learn to code and help nonprofits.
★ 38katom. atom 快捷键 shortcuts
★ 437vue-spa-template. The base code of vue.js project.
★ 662vue-router. 🚦 The official router for Vue 2
★ 19kgists. Gists for use in GraphGists.
★ 69