bit of an off topic post, but this demo includes barcode recognition within digital photos from an iPad:
Showing posts with label capture. Show all posts
Showing posts with label capture. Show all posts
Wednesday, July 11, 2012
Wednesday, May 9, 2012
OCR and PC Architecture
So just how important is your PC hardware when looking to use OCR Software? Many of the desktop products do not take advantage of multi-core CPUs, and can have laggard performance numbers when it comes to Optical Character Recognition, Intelligent Character Recognition and Optical Mark Recognition. Currently, playing with PSI:Capture, which offers a number of OCR options, and they have single, dual and quad core enablement in their licensing. Dual core runs about 1.7 times the speed, and quad core gives a 2.7x improvement.
Saturday, February 13, 2010
Why use OCR Software to perform full text conversion of images?
OCR Software
When we scan documents, they are just images, pictures of our paper. For many organizations, this scanned image is exactly what they need, and a little index information about the document is sufficient to provide them with retrieval capability.
So why take the time and spend the money to utilize OCR Software to convert the scanned document to a searchable format? Below are some reasons to always perform full text OCR of scanned documents:
When we scan documents, they are just images, pictures of our paper. For many organizations, this scanned image is exactly what they need, and a little index information about the document is sufficient to provide them with retrieval capability.
So why take the time and spend the money to utilize OCR Software to convert the scanned document to a searchable format? Below are some reasons to always perform full text OCR of scanned documents:
- Always provide every means possible for retrieval. Just using index fields to search for scanned documents may seem like a fantastic idea, but what if the document is misidentified? Or the indexer enters incorrect information? Performing a full text OCR of the document can provide an insurance policy that a document can always be found through full text search.
- Document Capture software today provides fast reliable OCR. Most capture software on the market provides the ability to automatically convert the documents to searchable format for a small expense. Some of the engines on the market can do the conversion at 100+ pages per minute, so there is really not much time wasted in the OCR conversion / recognition process.
- OCR to PDF for a format that contains both image and text in one container. Adobe provides the PDF image with hidden text option to give you a seachable file format that contains a pristine image.
- Plan for the worst case. Audits...legal issues...sometimes you need to search beyond the index fields, and full text can give you the ability to find the needle in the haystack.
Wednesday, December 30, 2009
Optical Character Recognition (OCR) and Capture
Optical Character Recognition (OCR) and Capture
So what is document capture software and what does it have to do with OCR applications. So, I think first, we need to differentiate between scanning software and capture software. Here is a good blog post that goes over the differences, with regards to SharePoint Scanning. Scanning Software just gives you the ability to convert paper to a digital form, and then OCR. Capture Software takes this a step further, and is really a catalyst for some enhanced processing with your recognition engine. Typical capture software will allow you to perform zone OCR, scan multiple documents in a single stack through separation, perform OCR based separation or even analyze the OCR text for expressions and then automatically extract the data. Document Capture software provides enhanced data extraction, as an example, as do other vendors like Kofax, AnyDoc, Captiva, etc.
So, I guess the whole point here is that OCR software in most cases just provides a basic framework for the conversion process. you really need a capture application to harness the true power of any OCR or recognition engine.
So what is document capture software and what does it have to do with OCR applications. So, I think first, we need to differentiate between scanning software and capture software. Here is a good blog post that goes over the differences, with regards to SharePoint Scanning. Scanning Software just gives you the ability to convert paper to a digital form, and then OCR. Capture Software takes this a step further, and is really a catalyst for some enhanced processing with your recognition engine. Typical capture software will allow you to perform zone OCR, scan multiple documents in a single stack through separation, perform OCR based separation or even analyze the OCR text for expressions and then automatically extract the data. Document Capture software provides enhanced data extraction, as an example, as do other vendors like Kofax, AnyDoc, Captiva, etc.
So, I guess the whole point here is that OCR software in most cases just provides a basic framework for the conversion process. you really need a capture application to harness the true power of any OCR or recognition engine.
Labels:
capture,
OCR Software,
Optical Character Recognition,
scanning
Subscribe to:
Posts (Atom)