Rare find

parsekit. Ruby document parsing toolkit with zero runtime dependencies. Parse PDFs, DOCX, XLSX, and images (with OCR) using a single, lightweight gem. Statically links MuPDF and Tesseract at compile time for hassle-free installation - no system libraries or external tools required.

github.com/scientist-labs/parsekit

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.