This is your work, valued
Freelance AI / ML Engineer
MambaViT. ViT architecture with Mamba instead of transformer backbone
★ 17knowledge_engine. Python
★ 3spatialfusion. This repository contains the package SpatialFusion, a deep-learning multimodal model for niche discovery in spatial transcritomics and histopathological data.
★ 40paseo. Orchestrate multiple coding agents from desktop and mobile
★ 12kbm25x. A fast, streaming-friendly BM25 search engine in Rust with mmap support
★ 54PaddleOCR. Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
★ 86klast30days-skill. AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthesizes a grounded summary
★ 55kpylate. Late Interaction Models Training & Retrieval
★ 876VectorChord. Scalable, fast, and disk-friendly vector search in Postgres, the successor of pgvecto.rs.
★ 1.8kautokernel. Autoresearch for GPU kernels. Give it any PyTorch model, go to sleep, wake up to optimized Triton kernels.
★ 1.5klangextract. A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
★ 38kpaper-qa. High accuracy RAG for answering questions from scientific documents with citations
★ 9kxtr-warp. XTR/WARP (SIGIR'25) is an extremely fast and accurate retrieval engine based on Stanford's ColBERTv2/PLAID and Google DeepMind's XTR.
★ 212dspydantic. DSPydantic: Auto-Optimize Your Prompts and Pydantic Models with DSPy
★ 328RADIO. Official repository for "AM-RADIO: Reduce All Domains Into One"
★ 1.9klightly. A python library for self-supervised learning on images.
★ 3.8kreasoning-from-scratch. Implement a reasoning LLM in PyTorch from scratch, step by step
★ 4.8kQwen3-VL-Embedding. Python
★ 1.3kjaxtyping. Type annotations and runtime checking for shape and dtype of JAX/NumPy/PyTorch/etc. arrays. https://docs.kidger.site/jaxtyping/
★ 1.8kUMIE_datasets. Python
★ 72a2ui. TypeScript
★ 16kProject-Imaging-X. Project Imaging-X: A Survey of 1000+ Open-Access Medical Imaging Datasets for Foundation Model Development
★ 468cutile-python. cuTile is a programming model for writing parallel kernels for NVIDIA GPUs
★ 2.1kstable-pretraining. Reliable, minimal and scalable library for pretraining foundation and world models
★ 296humanlayer. The best way to get AI coding agents to solve hard problems in complex codebases.
★ 11kopenvas-scanner. This repository contains the scanner component for Greenbone Community Edition.
★ 4.7ktemplates. Templates for ISO 13485, IEC 62304, ISO 14971 and IEC 62366 compliance.
★ 172cell-eval. Comprehensive suite for evaluating perturbation prediction models
★ 146mantis-viewer. Electron-based multiplexed imaging viewer
★ 27PufferLib. Puffing up reinforcement learning
★ 6.2kCopernicus-FM. Towards a Unified Copernicus Foundation Model for Earth Vision
★ 146inspect_ai. Inspect: A framework for large language model evaluations
★ 2.4kMRGen. [ICCV 2025] MRGen: Segmentation Data Engine for Underrepresented MRI Modalities
★ 41verifiers. Our library for RL environments + evals
★ 4.4kLLM-Trading-Lab. This repo powers my experiment where ChatGPT manages a real-money micro-cap stock portfolio.
★ 7.5kopenbench. Provider-agnostic, open-source evaluation infrastructure for language models
★ 795Archon. The first open-source harness builder for AI coding. Make AI coding deterministic and repeatable.
★ 23kclaude-code. Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
★ 140kaudio. Data manipulation and transformation for audio signal processing, powered by PyTorch
★ 2.9kLean. Lean Algorithmic Trading Engine by QuantConnect (Python, C#)
★ 21kCopilotKit. The Frontend Stack for Agents & Generative UI. React, Angular, Mobile, Slack, and more. Makers of the AG-UI Protocol
★ 36ktorchft. Fault tolerance for PyTorch (HSDP, LocalSGD, DiLoCo, Streaming DiLoCo)
★ 528gitingest. Replace 'hub' with 'ingest' in any GitHub URL to get a prompt-friendly extract of a codebase
★ 15kRareArena. A Comprehensive Rare Disease Diagnostic Dataset with nearly 50,000 patients covering more than 4000 diseases
★ 50ViT-Adapter. [ICLR 2023 Spotlight] Vision Transformer Adapter for Dense Predictions
★ 1.5kbeyond-nanogpt. Minimal and annotated implementations of key ideas from modern deep learning research.
★ 1.3kRadVLM. A Multitask Conversational Vision-Language Model for Radiology
★ 17MedVAE. [MIDL 2025] Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders
★ 195onlook. The Cursor for Designers • An Open-Source AI-First Design tool • Visually build, style, and edit your React App with AI
★ 26kames. AMES: Asymmetric and Memory-Efficient Similarity
★ 48VLM-Radiology-Agent-Framework. Jupyter Notebook
★ 220optuna. A hyperparameter optimization framework
★ 15kmedgemma. Jupyter Notebook
★ 1.6kmarker. Convert PDF to markdown + JSON quickly with high accuracy
★ 38kbatched. The Batched API provides a flexible and efficient way to process multiple requests in a batch, with a primary focus on dynamic batching of inference workloads.
★ 161embedding-atlas. Embedding Atlas is a tool that provides interactive visualizations for large embeddings. It allows you to visualize, cross-filter, and search embeddings and metadata.
★ 4.9kbaguetter. Baguetter is a flexible, efficient, and hackable search engine library implemented in Python. It's designed for quickly benchmarking, implementing, and testing new search methods. Baguetter supports sparse (traditional), dense (semantic), and hybrid retrieval methods.
★ 211deer-flow. An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
★ 78kCellTune-App. Public landing page and releases for CellTune.
★ 7nanoVLM. The simplest, fastest repository for training/finetuning small-sized VLMs.
★ 5kperception_models. State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!
★ 2.3ktextgrad. TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.
★ 3.7kpython-blosc2. A high-performance library for compressed ND arrays and columnar tables, with compute and indexing engines
★ 209c-blosc2. A fast, compressed, persistent binary data store library for C.
★ 580nnUNet. Python
★ 47kotaemon. An open-source RAG-based tool for chatting with your documents.
★ 26kpubmed_parser. :clipboard: A Python Parser for PubMed Open-Access XML Subset and MEDLINE XML Dataset
★ 735OpenManus. No fortress, purely open ground. OpenManus is Coming.
★ 58kai2-scholarqa-lib. Repo housing the open sourced code for the ai2 scholar qa app and also the corresponding library
★ 279SimDINO. [ICML 2025] Official Implementation for SimDINO/SimDINOv2
★ 206torchtitan. A PyTorch native platform for training generative AI models
★ 5.6khealthcareai-examples. Jupyter Notebook
★ 201fft. Python
★ 131deepcell-types. Python
★ 9olmocr. Toolkit for linearizing PDFs for LLM datasets/training
★ 19kwasp. The batteries-included full-stack framework for the AI era. Develop JS/TS web apps (React, Node.js, and Prisma) using declarative code that abstracts away complex full-stack features like auth, background jobs, RPC, email sending, end-to-end type safety, single-command deployment, and more.
★ 19kViDoRAG. [EMNLP 2025] ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents
★ 668siglip. Projects based on SigLIP (Zhai et. al, 2023) and Hugging Face transformers integration 🤗
★ 312conditional-flow-matching. TorchCFM: a Conditional Flow Matching library
★ 2.6ks1. s1: Simple test-time scaling
★ 6.7koumi. Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
★ 9.4kopen-r1. Fully open reproduction of DeepSeek-R1
★ 26kfiftyone. Refine high-quality datasets and visual AI models
★ 11kcolpali. The code used to train and run inference with the ColVision models, e.g. ColPali, ColQwen2, and ColSmol.
★ 2.7karctique. Python
★ 11virtues. A foundation model framework for analyzing multiplexed tissue imaging data across molecular, cellular and tissue scales enabling clinical diagnostics, biological discovery and patient case retrieval.
★ 78Merlin. [Nature 2026] Merlin is a 3D VLM for computed tomography that leverages both structured electronic health records (EHR) and unstructured radiology reports for pretraining.
★ 454CellViT-plus-plus. Python
★ 218markitdown. Python tool for converting files and office documents to Markdown.
★ 170kTransformerLens. A library for mechanistic interpretability of GPT-style language models
★ 3.7kpicotron. Minimalistic 4D-parallelism distributed training framework for education purpose
★ 2.3ktract. TrAct: Making First-layer Pre-Activations Trainable
★ 8jan. Jan is an open source alternative to ChatGPT that runs 100% offline on your computer.
★ 44ktorchSOM. Python
★ 1NanoPyx. Nanoscopy library for Python (NanoPyx, the successor to NanoJ) - focused on light microscopy and super-resolution imaging
★ 93smolagents. 🤗 smolagents: a barebones library for agents that think in code.
★ 29kdeepinv. DeepInverse: a PyTorch library for solving imaging inverse problems using deep learning
★ 787datasets. 🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
★ 22kuv. An extremely fast Python package and project manager, written in Rust.
★ 88kollama-python. Ollama Python library
★ 10kmmcv. OpenMMLab Computer Vision Foundation
★ 6.5kthe_well. A 15TB Collection of Physics Simulation Datasets
★ 4.3kVGen. Official repo for VGen: a holistic video generation ecosystem for video generation building on diffusion models
★ 3.2kml-aim. This repository provides the code and model checkpoints for AIMv1 and AIMv2 research projects.
★ 1.4kunsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kbyaldi. Use late-interaction multi-modal models such as ColPali in just a few lines of code.
★ 851trl. Train transformer language models with reinforcement learning.
★ 19krerankers. A lightweight, low-dependency, unified API to use all common reranking and cross-encoder models.
★ 1.6kCurator. Scalable data pre processing and curation toolkit for LLMs
★ 1.7kragflow. RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
★ 86ktorchmetrics. Machine learning metrics for distributed, scalable PyTorch applications.
★ 2.5kssl-data-curation. PyTorch code for hierarchical k-means -- a data curation method for self-supervised learning
★ 246UCE. UCE is a zero-shot foundation model for single-cell gene expression data
★ 319Python. All Algorithms implemented in Python
★ 223kscenic. Scenic: A Jax Library for Computer Vision Research and Beyond
★ 3.8kcellSAM. Codebase for "A Foundation Model for Cell Segmentation"
★ 208sam2. The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 20kneurallambda. Reasoning Computers. Lambda Calculus, Fully Differentiable. Also Neural Stacks, Queues, Arrays, Lists, Trees, and Latches.
★ 289storm. An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.
★ 30kPEERCE. PEERCE, PD-L1 Expression Estimation for Rare Cancer Entities
★ 4ell. A language model programming library.
★ 5.9khover_next_inference. Inference code for HoVer-NeXt
★ 51OpenHands. 🙌 OpenHands: AI-Driven Development
★ 83kMiniCPM-V. A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
★ 26kInternVL. [CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
★ 10kdspy. DSPy: The framework for programming—not prompting—language models
★ 36kNimbus-Inference. Python
★ 43openvino. OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
★ 11kOpenSearch. 🔎 Open source distributed and RESTful search engine.
★ 13kmagiclens. [ICML'24 Oral] "MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions"
★ 211direct. Deep learning framework for MRI reconstruction
★ 308UnSAM. [NeurIPS 2024] Code release for "Segment Anything without Supervision"
★ 503openst. Open-ST: profile and analyze tissue transcriptomes in 3D with high resolution in your lab
★ 116Dragonfly. Python
★ 81performance. :muscle: Models' quality and performance metrics (R2, ICC, LOO, AIC, BF, ...)
★ 1.1kVane. Vane is an AI-powered answering engine.
★ 36kautogen. A programming framework for agentic AI
★ 60kChatWizard. OpenAI chat client desktop app (Windows, MacOS, Linux)
★ 177Bend. A massively parallel, high-level programming language
★ 20kX2-VLM. All-In-One VLM: Image + Video + Transfer to Other Languages / Domains (TPAMI 2023)
★ 169ThunderKittens. Tile primitives for speedy kernels
★ 3.6kOpenCRISPR. AI-generated gene editing systems
★ 1.2khover_next_train. Training/Evaluation code for HoVer-NeXt
★ 18SynthSR. A framework for joint super-resolution and image synthesis, without requiring real training data
★ 192multimodal. TorchMultimodal is a PyTorch library for training state-of-the-art multimodal multi-task models at scale.
★ 1.7kaxolotl. Go ahead and axolotl questions
★ 12kvitessce. Vitessce is a visual integration tool for exploration of spatial single-cell experiments.
★ 261LLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 25kscaling_on_scales. When do we not need larger vision models?
★ 419uvadlc_notebooks. Repository of Jupyter notebook tutorials for teaching the Deep Learning Course at the University of Amsterdam (MSc AI), Fall 2023
★ 3.2kdatatrove. Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.
★ 3.2ksubobjects. Official repository of paper "Subobject-level Image Tokenization" (ICML-25)
★ 94Me-LLaMA. A novel medical large language model family with 13/70B parameters, which have SOTA performances on various medical tasks
★ 166professional-programming. A collection of learning resources for curious software engineers
★ 51kfast-soft-sort. Fast Differentiable Sorting and Ranking
★ 624torchreg. Lightweight image registration library using PyTorch
★ 28cookiecutter-napari-plugin. Cookiecutter for napari plugins
★ 73datasets. Images and other data from the JUMP Cell Painting Consortium
★ 187dynamic-network-architectures. Python
★ 158ivy. Convert Machine Learning Code Between Frameworks
★ 14kcatalyst. Accelerated deep learning R&D
★ 3.4kmasif. MaSIF- Molecular surface interaction fingerprints. Geometric deep learning to decipher patterns in molecular surfaces.
★ 767RP3D-Diag. Code implementation of RP3D-Diag
★ 17cookiecutter-scverse. Cookiecutter template for scverse
★ 90cosmopolitan. build-once run-anywhere c library
★ 21kspatialdata. An open and interoperable data framework for spatial omics data
★ 383vstar. PyTorch Implementation of "V* : Guided Visual Search as a Core Mechanism in Multimodal LLMs"
★ 707PySyft. Perform data science on data that remains in someone else's server
★ 9.9kdeepcell-label. Cloud-based data annotation tools for biological images
★ 85fpng-python. Python bindings for fpng
★ 192024_Chandrasekaran_NatureMethods_CPJUMP1. Jupyter Notebook
★ 72mamba. Mamba SSM architecture
★ 19kNimbus. Python
★ 36DToP. Boosting vision transformers for image retrieval, proposed design of Deep Token Pooling(DToP)
★ 38build-your-own-x. Master programming by recreating your favorite technologies from scratch.
★ 533kCellViT. CellViT: Vision Transformers for Precise Cell Segmentation and Classification
★ 391triton. Development repository for the Triton language and compiler
★ 20kPMC-VQA. PMC-VQA is a large-scale medical visual question-answering dataset, which contains 227k VQA pairs of 149k images that cover various modalities or diseases.
★ 236UniBrain. An offcial implementation for UniBrain: Universal Brain MRI Diagnosis with Hierarchical Knowledge-enhanced Pre-training
★ 39cucim. cuCIM - RAPIDS GPU-accelerated image processing library
★ 466pytorch-stardist. PyTorch implementation of 2D and 3D StarDist - Object Detection with Star-convex Shapes
★ 29wsinfer. 🔥 🚀 Blazingly fast pipeline for patch-based classification in whole slide images
★ 86tuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kchromatix. Differentiable wave optics using JAX! Documentation can be found at https://chromatix.readthedocs.io
★ 209big_vision. Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.
★ 3.5kcellulus. Unsupervised Instance Segmentation in Microscopy
★ 27coding-interview-university. A complete computer science study plan to become a software engineer.
★ 357kunicom. Large-Scale Visual Representation Model
★ 700ImPartial. Interactive deep learning cell segmentation using partial annotations [MICCAI'26]
★ 45aydin. Aydin — User-friendly, Fast, Self-Supervised Image Denoising for All.
★ 174x-transformers. A concise but complete full-attention transformer with a set of promising experimental features from various papers
★ 5.9kjax. Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
★ 36kcandle. Minimalist ML framework for Rust
★ 21kopen-metric-learning. Metric learning and retrieval pipelines, models and zoo.
★ 995med-flamingo. Python
★ 452LLaVA-Med. Large Language-and-Vision Assistant for Biomedicine, built towards multimodal GPT-4 level capabilities.
★ 2.2kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kark-analysis. Integrated pipeline for multiplexed image analysis
★ 107Cell_ACDC. A Python GUI-based framework for segmentation, tracking and cell cycle annotations of microscopy data
★ 200GenerativeModels. MONAI Generative Models makes it easy to train, evaluate, and deploy generative models and related applications
★ 764Open-Assistant. OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
★ 37kRadImageGAN.
★ 26