This is your work, valued
Awesome-Captioning. A curated list of Multimodal Captioning related research(including image captioning, video captioning, and text captioning)
★ 113RNABenchmark. [NeurIPS 2024] BEACON: Benchmark for Comprehensive RNA Tasks and Language Models
★ 64Genomics-FM. Genomics-FM
★ 13Awesome-Multimodal-Pretraining.
★ 7Cross-MolecularBenchmark. Source code of "A Comprehensive Cross-Molecular Benchmark for Language Model Evaluation and Tasks in Biological Sequence Understanding".
★ 3CENO. Python
★ 8ceno.github.io. HTML
★ 2smp-ppi. Online server: https://zhanggroup.org/SMP-PPI
★ 4smp-docking. Online server: https://zhanggroup.org/SMP-docking
★ 4smp-contact. Online server: https://zhanggroup.org/SMP-contact
★ 4Awesome-Scientific-Datasets-and-LLMs. A curated collection of papers, datasets, and resources on Scientific Datasets and Large Language Models (LLMs)
★ 458Intern-S1. A Scientific Multimodal Foundation Model
★ 842alphagenome. This API provides programmatic access to the AlphaGenome model developed by Google DeepMind.
★ 2kCross-MolecularBenchmark. Source code of "A Comprehensive Cross-Molecular Benchmark for Language Model Evaluation and Tasks in Biological Sequence Understanding".
★ 3golitex. Litex: The Language Where Mathematics Verifies Itself.
★ 659audio-ai-hub. The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
★ 949ProDMM. Source code of ProDMM. The paper is titled with "Unveiling Protein-DNA Interdependency: Harnessing Unified Multimodal Sequence Modeling, Understanding and Generation".
★ 10BiRefNet. [CAAI AIR'24] Bilateral Reference for High-Resolution Dichotomous Image Segmentation
★ 4kGenomics-FM. Genomics-FM
★ 13chai-lab. Chai-1, SOTA model for biomolecular structure prediction
★ 2kLucaOne. The resources of LucaOne, including: the model code, training scripts, embedding inference code, and trained checkpoints.
★ 367esm. Jupyter Notebook
★ 2.9kLucaOneTasks. The project of the downstream tasks based on LucaOne's Embedding.
★ 77RNABenchmark. [NeurIPS 2024] BEACON: Benchmark for Comprehensive RNA Tasks and Language Models
★ 64evo. Biological foundation modeling from molecular to genome scale
★ 1.5kunilm. Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22kImageBind. ImageBind One Embedding Space to Bind Them All
★ 9.1kAudio-visual-sound-localization. Audio-visual sound localization
★ 11ER-SAN. Implementation of our IJCAI2022 oral paper, ER-SAN: Enhanced-Adaptive Relation Self-Attention Network for Image Captioning.
★ 25notears. DAGs with NO TEARS: Continuous Optimization for Structure Learning
★ 683VolumetricCapture. A multi-sensor capture system for free viewpoint video.
★ 530yolov5. Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
★ 58kOmegaFold. OmegaFold Release Code
★ 625mmdetection3d. OpenMMLab's next-generation platform for general 3D object detection.
★ 6.5kBalancedMSE. [CVPR 2022 Oral] Balanced MSE for Imbalanced Visual Regression https://arxiv.org/abs/2203.16427
★ 394DIPS-Plus. The Enhanced Database of Interacting Protein Structures for Interface Prediction
★ 55Awesome-CLIP. Awesome list for research on CLIP (Contrastive Language-Image Pre-Training).
★ 1.2kClipCap-Chinese. 基于ClipCap的看图说话Image Caption模型
★ 325DeepInteract. A geometric deep learning framework (Geometric Transformers) for predicting protein interface contacts. (ICLR 2022)
★ 66hh-suite. Remote protein homology detection suite.
★ 627Awesome-Multimodal-Pretraining.
★ 7Awesome-Transformer-Attention. An ultimately comprehensive paper list of Vision Transformer/Attention, including papers, codes, and related websites
★ 5.1kPaper-Writing-Tips. MLNLP社区用来帮助大家避免论文投稿小错误的整理仓库。 Paper Writing Tips
★ 4.6ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kesm. Evolutionary Scale Modeling (esm): Pretrained language models for proteins
★ 4.2kOntoProtein. [ICLR 2022] OntoProtein: Protein Pretraining With Gene Ontology Embedding
★ 152CBTrans. Python
★ 24HowToCook. Programmer's guide about how to cook at home.
★ 101kABINet-PP. ABINet++: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Spotting
★ 90Awesome-Captioning. A curated list of Multimodal Captioning related research(including image captioning, video captioning, and text captioning)
★ 113ColabFold. Making Protein folding accessible to all!
★ 2.9kUni-Fold. Python
★ 93ABINet. Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition
★ 464Talking-Face_PC-AVS. Code for Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation (CVPR 2021)
★ 959mmf. A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)
★ 5.6kmeshed-memory-transformer. Meshed-Memory Transformer for Image Captioning. CVPR 2020
★ 546grid-feats-vqa. Grid features pre-training code for visual question answering
★ 269collaborative-experts. Video embeddings for retrieval with natural language queries
★ 344Oscar. Oscar and VinVL
★ 1.1kDiscriminative-Sounding-Objects-Localization. Code for Discriminative Sounding Objects Localization (NeurIPS 2020)
★ 61neuraltalk2. Efficient Image Captioning code in Torch, runs on GPU
★ 5.6kneuraltalk. NeuralTalk is a Python+numpy project for learning Multimodal Recurrent Neural Networks that describe images with sentences.
★ 5.5kawesome-audio-visual. A curated list of different papers and datasets in various areas of audio-visual processing
★ 775Revisit-MMT. Python
★ 25Awesome-Visual-Captioning. This repository focus on Image Captioning & Video Captioning & Seq-to-Seq Learning & NLP
★ 3self-critical.pytorch. Unofficial pytorch implementation for Self-critical Sequence Training for Image Captioning. and others.
★ 1kfairseq-zh-en. NMT for chinese-english using fairseq
★ 213DeepRL. Deep Reinforcement Learning Lab, a platform designed to make DRL technology and fun for everyone
★ 2.6k