Rare find

text-extract-api. Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown

github.com/CatchTheTornado/text-extract-api

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.