20 projects tagged "OCR"

No download Website Updated 20 May 2014 Solr-Connector-Files

Screenshot
Pop 21.05
Vit 3.13

Solr-Connector-Files crawls and indexes directories and files from your filesystem (whatever is mountable to Linux) into Apache Solr. It features extraction of file contents with Tika, which extracts metadata and text form many document and file formats. It also integrates automatic text recognition (OCR) for images, photos, and PDFs using Tesseract OCR.

No download Website Updated 20 Mar 2014 Pyocr

Screenshot
Pop 149.35
Vit 3.50

Pyocr is a simple Python wrapper for OCR engines (Tesseract, Cuneiform, etc.). It supports Python 2.7 and Python 3.x, and requires Pillow.

Download Website Updated 05 Mar 2014 Paperwork

Screenshot
Pop 193.66
Vit 2.40

Paperwork is a GUI to make papers easily searchable using OCR. The basic idea behind Paperwork is "scan & forget" : You should be able to just scan a new document and forget about it until the day you need it again.

Download Website Updated 29 Apr 2014 GlyphViewer

Screenshot
Pop 24.52
Vit 18.67

GlyphViewer builds translations in a multitude of modern languages from text in your images (and even ancient writing) using advanced OCR technology and online machine translators. This way, you will not only improve the SEO rating of your images, but your online content can be understood by your users in their native language.

No download No website Updated 28 May 2014 Character Recognition

Screenshot
Pop 301.21
Vit 92.80

Character Recognition is an Android app that allows the user to take a photo (or use existing image files on the device) and then apply the Tesseract OCR engine to extract the text in the photo. It is currently supporting English text, but other language support will be added in the future.

No download Website Updated 14 Oct 2013 getxbook

Screenshot
Pop 114.89
Vit 5.33

getxbook is a collection of tools to download books from websites. There are tools to download from Google Books' "book preview", Amazon's "look inside the book", and Barnes and Noble's "book viewer". There is an optional GUI written in Tcl/Tk, and some shell scripts using OCR to create plain text or searchable PDFs and DjVu files from the downloaded books.

Download No website Updated 19 Nov 2011 Lector

Screenshot
Pop 31.08
Vit 1.00

Lector can help you scan your papers and create text documents. It lets you select areas on which you want to use OCR. Then you can run tesseract-ocr simply by clicking a button. The resulting text can be proofread, formatted, and edited directly in Lector.

Download No website Updated 26 Apr 2014 Aspose.OCR for .NET

Screenshot
Pop 69.30
Vit 10.99

Aspose.OCR for .NET is a character recognition component built to allow developers to add OCR functionality in their ASP .NET Web applications, Web services, and applications. It provides a simple set of classes for controlling character recognition tasks and supports BMP and TIFF.

Download Website Updated 19 Jun 2012 MALODOS

Screenshot
Pop 77.20
Vit 3.60

MALODOS helps you to scan, store, and easily retrieve all your personal documents. Its storage format is open and documented, so your document archive can remain accessible even without MALODOS. The documents themselves are stored as standard PDF files, while their metadata (such as title, tags, and description) are stored into a separate SQLite database in an open format. With MALODOS, you can also manage existing files in PDF, JPEG, TIFF, and other formats, so you can still use the documents that you've already scanned. You can connect to any external OCR program to give access to a fulltext search feature.

No download No website Updated 04 Mar 2011 OCR2DATA

Screenshot
Pop 19.44
Vit 34.39

OCR2DATA is a full OCR stack for document digitization analysis and OCR. It provides external connection by way of an API, standard document exchange formats, and a database.

Screenshot

Project Spotlight

cclite

LETS and community currency software.

Screenshot

Project Spotlight

bind

Berkeley Internet Name Domain