This is your work, valued
Building voice, vision, and GPU infrastructure nobody asked for. Ships fast, documents never. This is AI.
Qwen3-TTS-finetune. One-command fine-tuning for Qwen3-TTS text-to-speech model with custom voice samples
27gemsAPI. A FastAPI and MCP server connected to Supabase table for managing, testing, and programmatically accessing LLM System Instructions, System Prompts, BOTs, Google Gems.
12ClaudeWebUI-Docker. Dockerized Claude Code UI - Web interface for Claude Code CLI with Docker deployment support
10RMBanana. Dockerized Web UI for removing invisible AI watermarks from Google Gemini-generated images using reverse alpha blending
8InfiniteTalk-Google-Collab. Jupyter Notebook
7canary-qwen-2.5b-RunPod. NVIDIA NeMo Canary-Qwen-2.5B is an English speech recognition model that achieves state-of-the art performance on multiple English speech benchmarks.
5higgs-audio-v2. a powerful audio foundation model pretrained on over 10 million hours of audio data and a diverse set of text data
5SeedVR-RunPod. 🎬 SeedVR RunPod Container - Lightweight Docker container for SeedVR video restoration
3astral. Organize your GitHub Stars with ease - A powerful Laravel + Vue.js application for managing starred repositories with tags, notes, and smart filters
3VoxCPM-Runpod. Runpod for executing this TTS Model https://github.com/OpenBMB/VoxCPM/
3zimage-serverless. A serverless for Runpod that will run inference using ZImage model and will take LoRA as input.
2crawl-mcp. A comprehensive Model Context Protocol (MCP) server that wraps the powerful crawl4ai library.
2VibeVoice-finetune-easy. Simplified scripts for fine-tuning VibeVoice speech synthesis models with LoRA. Painless fine-tuning with reasonable defaults, supporting both local GPU and Google Colab workflows.
1llama.cpp-serverless. A runpod serverless that runs llama.cpp
1Moss-TTS-Runpod. Scalable Text-to-Speech inference server powered by OpenMOSS-Team/MOSS-TTS on RunPod Serverless GPU infrastructure. Features voice cloning, continuation mode, S3 uploads, and lazy model loading.
1chatterbox-runpod-serverless. ChatterBox Runpod Serverless: Zero-shot voice cloning TTS with 23+ languages
1OmniTry-Runpod. A project to containerize OmniTry to run on Runpod as a stand alone container.
1addit-Runpod. a training-free approach that extends diffusion models' attention mechanisms to incorporate information from three key sources: the scene image, the text prompt, and the generated image itself
1parakeet-runpod. A runpod serverless implementing Nvidia's Multilingual Speech-to-Text Model
1runhub. AI-powered batch image generation interface using Google Gemini for prompt engineering and custom python pipelines for generating images using Flux.2-klein-9b and zImage base models. Also supports custom workflows from Running hub. Enhance and upscale options. . Built with SvelteKit 5 + Docker.
1spark-tts-runpod-serverless. SparkTTS runpod serverless
1Musabi-Runpod. Runpod template for Musabi-Tuner and Z-Image Base Lora Training.
1MTVCraft-RunPod. An Open Veo3-style Audio-Video Generation
1VibeVoice-Serverless. Production-ready RunPod serverless deployment of VibeVoice 7B with voice cloning
1