Rare find

imagecat. ImageCat is an Apache OODT RADIX application that uses Apache Solr, Apache Tika and Apache OODT to ingest 10s of millions of files (images,but could be extended to other files) in place, and to extract metadata and OCR information from those files/images using Tika and Tesseract OCR.

github.com/chrismattmann/imagecat

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.