Solved

Scanned pages into pdf

Posted on 2006-06-27
3
314 Views
Last Modified: 2010-04-17
hello all does anyone  know how are scanned pages containing text converted into pdf and what is the format in which text is stored in the pdf. Can this text be directly extracted from the pdf or some technique like OCR is required to be used.

Thanks
0
Comment
Question by:jhav1594
[X]
Welcome to Experts Exchange

Add your voice to the tech community where 5M+ people just like you are talking about what matters.

  • Help others & share knowledge
  • Earn cash & points
  • Learn & ask questions
3 Comments
 
LVL 1

Accepted Solution

by:
jm021196 earned 500 total points
ID: 16994505
It reallly depends on what app you are using and the quality of the page.

If the PDF Converting program which is being used to take the image from the scanner can recognise the text as text then its stored as text in the PDF File.

If the converting program cannot recognise it as text then it gets saved in a variety of image formats depending on which one suits it best. There really is no way to tell how its saved in advance.

PDF Files use a combination of vector, raster and text formates to give the best compression and viewability and so converting to PDF is a very difficult thing to undo... especiall if its not possible to tell in advance if its going to be in text or not.

I would suggest that a OCR system is the best way forward.

Thanks
mitch
0

Featured Post

The Ultimate Checklist to Optimize Your Website

Websites are getting bigger and complicated by the day. Video, images, custom fonts are all great for showcasing your product/service. But the price to pay in terms of reduced page load times and ultimately, decreased sales, can lead to some difficult decisions about what to cut.

Question has a verified solution.

If you are experiencing a similar issue, please ask a related question

This article is meant to give a basic understanding of how to use R Sweave as a way to merge LaTeX and R code seamlessly into one presentable document.
Displaying an arrayList in a listView using the default adapter is rarely the best solution. To get full control of your display data, and to be able to refresh it after editing, requires the use of a custom adapter.
Simple Linear Regression
Introduction to Processes

726 members asked questions and received personalized solutions in the past 7 days.

Join the community of 500,000 technology professionals and ask your questions.

Join & Ask a Question