This is your work, valued
Dr in Computer Science, Data Scientist, Multimedia Engineer - University of Alicante. Mainly focused in Optical Music Recognition research.
SMT. Official implementation of the Sheet Music Transformer
★ 80SMT-plusplus. Official implementation of the Sheet Music Transformer ++
★ 33ijdar-e2e-pianoform. Source code for the paper "End-to-end optical music recognition for pianoform sheet music"
★ 6TheRookieThiefCPC. Amstrad CPC 464 assembly language videogame developed for the CPCRetroDev 2018.
★ 2icdar-2023-amnlt. Source code for the paper "A Holistic Approach for Aligned Music Notation and Lyrics Transcription"
★ 2transformers_omr. Python
★ 2ismir_fpomr. Source code for the paper "End-To-End Full-Page Optical Music Recognition For Mensural Notation"
★ 1omr_two_dim_exp. Experiment for ICFHR to comapre Seq 2 Seq, Seq 2 Seq + Attention and CTC for staff symbols classification
★ 1img-ed-det. Python
★ 1JoustGame. Classic Joust game with SFML and C++ for the subject Videogame Fundamentals at the University of Alicante.
★ 1fmod-example-ecs. C++
★ 1omr_img_krn. Python
★ 1DCAGII-Game-Template. Template para el desarrollo de un juego básico con C++
★ 1Inspire. Linux PC video game made from scratch with C++ and OpenGL by Zenon Games
★ 1MuSViT. ECCV 2026 publication: "MuSViT: A Foundation Vision Model\\for Sheet Music Representation" oficial repository
★ 56PaddleOCR. Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
★ 87kGLM-OCR. GLM-OCR: Accurate × Fast × Comprehensive
★ 7.2kpalmier-pro. macOS video editor built for AI
★ 13kcrawl4ai. 🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
★ 76kOverleafMCP. Model Context Protocol (MCP) server that lets AI assistants read Overleaf projects, parse LaTeX document structure, and push section-level edits back via Git. Compatible with Claude Desktop, Cursor, Windsurf, and any MCP client.
★ 214legalize-pipeline. Legislación española como repositorio Git — cada ley es un fichero Markdown, cada reforma un commit
★ 48unsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kautoskills. One command. Your entire AI skill stack. Installed.
★ 6.6kpaperclip. The open-source app everyone uses to manage agents at work
★ 75kgsap-skills. Official AI skills for GSAP. These skills teach AI coding agents how to correctly use GSAP (GreenSock Animation Platform), including best practices, common animation patterns, and plugin usage.
★ 13klegalize-es. Legislación de España en Markdown, versionada como git. Cada ley es un archivo, cada reforma un commit.
★ 1.9kclaude-howto. A visual, example-driven guide to Claude Code — from basic concepts to advanced agents, with copy-paste templates that bring immediate value.
★ 41kmarketingskills. Marketing skills for Claude Code and AI agents. CRO, copywriting, SEO, analytics, and growth engineering.
★ 42kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74klegato. Official codebase for paper "LEGATO: Large-Scale End-to-End Generalizable Approach to Typeset OMR".
★ 65gstack. Use Garry Tan's exact Claude Code setup: 23 opinionated tools that serve as CEO, Designer, Eng Manager, Release Manager, Doc Engineer, and QA
★ 125kverifactu. Librería TypeScript production-ready para integrar con el sistema Verifactu de la AEAT (Agencia Tributaria Española).
★ 5assistant-ui. Typescript/React Library for AI Chat💬🚀
★ 11kxberg. A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 98+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
★ 8.7kDeepTutor. DeepTutor: Lifelong Personalized Tutoring. https://deeptutor.info/.
★ 31ktextual. The lean application framework for Python. Build sophisticated user interfaces with a simple Python API. Run your apps in the terminal and a web browser.
★ 37kratatui. A Rust crate for cooking up terminal user interfaces (TUIs) 👨🍳🐀 https://ratatui.rs
★ 22kcyclopts. Intuitive, easy CLIs based on python type hints.
★ 1.2kdocling. Get your documents ready for gen AI
★ 64kImReflect. Reflection-based ImGui wrapper
★ 345cpp-httplib. A C++ header-only HTTP/HTTPS server and client library
★ 17ktoon. 🎒 Token-Oriented Object Notation (TOON) – compact, human-readable serialization of JSON data for LLM prompts. TypeScript SDK, CLI, benchmarks.
★ 25kGameBoyTemplate. Project template for Game Boy / GBC games written using GBDK.
★ 27Architecture. An example of how I like to architect applications in C++
★ 257i18n_keyval. Easy to use and customizable C++ internationalization library.
★ 11system-prompts-and-models-of-ai-tools. FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Trae, Traycer AI, VSCode Agent, Warp.dev, Windsurf, Xcode, Z.ai Code, Dia & v0. (And other Open Sourced) System Prompts, Internal Tools & AI Models
★ 142khydra. Hydra is a framework for elegantly configuring complex applications
★ 11kOpenCut. The open-source CapCut alternative
★ 80kOpenRLHF. An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9.9kPyTorch-Scratch-Vision-Transformer-ViT. Simple and easy to understand PyTorch implementation of Vision Transformer (ViT) from scratch, with detailed steps. Tested on common datasets like MNIST, CIFAR10, and more.
★ 157DAN. Python
★ 12FasterDAN. Python
★ 4vexflow. A JavaScript library for rendering music notation and guitar tablature.
★ 4.4kdaniel. This repository contain the implementation of DANIEL. (A fast Document Attention Network for Information Extraction and Labeling of handwritten documents)
★ 22deer-flow. An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
★ 78kppo. An implementation of PPO in Pytorch
★ 126META-DAN. Python
★ 11slidev. Presentation Slides for Developers
★ 48koumi. Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
★ 9.4kvlms-zero-to-hero. This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.
★ 1.2krecommenders. Best Practices on Recommendation Systems
★ 22kannotated-mamba. Annotated version of the Mamba paper
★ 501annotated-s4. Implementation of https://srush.github.io/annotated-s4
★ 519bolt.diy. Prompt, run, edit, and deploy full-stack web applications using any LLM you want!
★ 20kmarkitdown. Python tool for converting files and office documents to Markdown.
★ 170kghostty. 👻 Ghostty is a fast, feature-rich, and cross-platform terminal emulator that uses platform-native UI and GPU acceleration.
★ 59kyt-dlp. A feature-rich command-line audio/video downloader
★ 181kawesome-styled-components. A curated list of awesome styled-components resources 💅
★ 3.5kmamba. Pytorch (Lightning) implementation of the Mamba model
★ 38uv. An extremely fast Python package and project manager, written in Rust.
★ 88kMessiScriptInterpreter. Intérprete de MessiScript, un lenguaje de programación esotérico en el que cada código es una jugada de Messi.
★ 297manim. A community-maintained Python framework for creating mathematical animations.
★ 40kmanim. Animation engine for explanatory math videos
★ 89kpytorch-attention. 🦖Pytorch implementation of popular Attention Mechanisms, Vision Transformers, MLP-Like models and CNNs.🔥🔥🔥
★ 545flamingo-pytorch. Implementation of 🦩 Flamingo, state-of-the-art few-shot visual question answering attention net out of Deepmind, in Pytorch
★ 1.3kOMR-Datasets. Collection of datasets used for Optical Music Recognition
★ 390kernpy. Python package that provides comprehensive tools for working with symbolic modern and mensural notations in Humdrum format. kernpy is a fully open-source project open to contributions.
★ 15void. TypeScript
★ 29ktypst. A markup-based typesetting system that is powerful and easy to learn.
★ 55kattentions. PyTorch implementation of some attentions for Deep Learning Researchers.
★ 548LoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kLLM101n. LLM101n: Let's build a Storyteller
★ 38kminGPT. A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training
★ 25kperceiver-pytorch. Implementation of Perceiver, General Perception with Iterative Attention, in Pytorch
★ 1.2kcoolify. An open-source, self-hostable PaaS alternative to Vercel, Heroku & Netlify that lets you easily deploy static sites, databases, full-stack applications and 280+ one-click services on your own servers.
★ 60kplaywright-python. Python version of the Playwright testing and automation library.
★ 15kAwesome-state-space-models. Collection of papers on state-space models
★ 620Vane. Vane is an AI-powered answering engine.
★ 36kbuild-nanogpt. Video+code lecture on building nanoGPT from scratch
★ 5.4kstreamlit. Streamlit — A faster way to build and share data apps.
★ 45kpix2seq. Pix2Seq codebase: multi-tasks with generative modeling (autoregressive and diffusion)
★ 945mamba-tiny. Simple, minimal implementation of the Mamba SSM in one pytorch file. Using logcumsumexp (Heisen sequence).
★ 134converter21. A music21-extending set of converters (Humdrum and MEI readers and writers), as well as a CLI app to convert music notation file formats
★ 30awesome-test-time-adaptation. Collection of awesome test-time (domain/batch/instance) adaptation methods
★ 1.3kzeta. Build high-performance AI models with modular building blocks
★ 598Gemini. The open source implementation of Gemini, the model that will "eclipse ChatGPT" by Google
★ 466LLMs-from-scratch. Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
★ 100ktypstudio. A W.I.P desktop application for a new typesetting language, typst.
★ 735transformer-debugger. Python
★ 4.1kzed. Code at the speed of thought – Zed is a high-performance, multiplayer code editor from the creators of Atom and Tree-sitter.
★ 88ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163koemer. End-to-end Optical Music Recognition (OMR) system. Transcribe phone-taken music sheet image into MusicXML, which can be edited and converted to MIDI.
★ 779a2s-transformer. A Transformer approach for polyphonic Audio-to-Score (A2S) transcription (ICASSP 2024)
★ 15ConvNeXt. Code release for ConvNeXt model
★ 6.4kopen_flamingo. An open-source framework for training large multimodal models.
★ 4.1kEmotiVoice. EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
★ 8.5kSmashcima. Training data synthesizer for OMR
★ 13musicdiff. A music score notation diff package (and command line tool)
★ 64Mashcima. Handwritten music image synthesizer for HMR
★ 21MeasureDetector. A Deep Learning based detector for measures in musical scores
★ 50AgentVerse. 🤖 AgentVerse 🪐 is designed to facilitate the deployment of multiple LLM-based agents in various applications, which primarily provides two frameworks: task-solving and simulation
★ 5.1kChatDev. ChatDev 2.0: Dev All through LLM-powered Multi-Agent Collaboration
★ 34kbuild-your-own-x. Master programming by recreating your favorite technologies from scratch.
★ 533kLazyVim. Neovim config for the lazy
★ 27kDALLE-pytorch. Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
★ 5.6kmlm-pytorch. An implementation of masked language modeling for Pytorch, made as concise and simple as possible
★ 181vector-quantize-pytorch. Vector (and Scalar) Quantization, in Pytorch
★ 4kmusiclm-pytorch. Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch
★ 3.3kopeninterpreter. A coding agent for open models like Kimi K3
★ 67kcs-video-courses. List of Computer Science courses with video lectures.
★ 83kML-Papers-Explained. Explanation to key concepts in ML
★ 8.6khumdrum-data. Humdrum data files download interface.
★ 49RetNet. Huggingface compatible implementation of RetNet (Retentive Networks, https://arxiv.org/pdf/2307.08621.pdf) including parallel, recurrent, and chunkwise forward.
★ 227WordStylist. Official PyTorch Implementation of "WordStylist: Styled Verbatim Handwritten Text Generation with Latent Diffusion Models" - ICDAR 2023
★ 83scholarphi. An interactive PDF reader.
★ 429unilm. Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
★ 22kMindmap. This repository will contain many mindmaps for cyber security technologies, methodologies, courses, and certifications in a tree structure to give brief details about them
★ 9.2kaudiocraft. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
★ 24kVATr. Python
★ 89synthtiger. Official Implementation of SynthTIGER (Synthetic Text Image Generator), ICDAR 2021
★ 579jrp-scores. Digital scores for all composers in the Josquin Research Project, which focuses on vocal music, ca. 1420–1520.
★ 25segment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55kSwinDocSegmenter. [ICDAR 2023] (Oral) An End-to-End Unified Domain Adaptive Transformer for Document Instance Segmentation
★ 74video2x. A machine learning-based video super resolution and frame interpolation framework. Est. Hack the Valley II, 2018.
★ 21kAwesome-Transformer-Attention. An ultimately comprehensive paper list of Vision Transformer/Attention, including papers, codes, and related websites
★ 5.1kdessurt. Official implementation for Dessurt: Document end-to-end self-supervised understanding and recognition transformer
★ 62SparK. [ICLR'23 Spotlight🔥] The first successful BERT/MAE-style pretraining on any convolutional network; Pytorch impl. of "Designing BERT for Convolutional Networks: Sparse and Hierarchical Masked Modeling"
★ 1.4kpytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kvit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kInstagramUnfollowers. Check if people follows you back on Instagram.
★ 4.7kcrnn_seq2seq_ocr_pytorch. Extremely simple implement for Chinese OCR by PyTorch.
★ 346PlotNeuralNet. Latex code for making neural networks diagrams
★ 25k3d-game-shaders-for-beginners. 🎮 A step-by-step guide to implementing SSAO, depth of field, lighting, normal mapping, and more for your 3D game.
★ 20kScoreGenerator. Python Scores Generator
★ 2transformer. A TensorFlow Implementation of the Transformer: Attention Is All You Need
★ 4.5kattention-is-all-you-need-pytorch. A PyTorch implementation of the Transformer model in "Attention is All You Need".
★ 9.8k