Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.debian.user > #185963

Re: xsane & tesseract

From Joe Pfeiffer <pfeiffer@cs.nmsu.edu>
Newsgroups linux.debian.user
Subject Re: xsane & tesseract
Date 2017-08-26 06:30 +0200
Message-ID <uiATn-4Zy-1@gated-at.bofh.it> (permalink)
References <uiyox-3f2-3@gated-at.bofh.it> <uiz1f-3Ni-3@gated-at.bofh.it>
Organization A noiseless patient Spider

Show all headers | View raw


Doug <dmcgarrett@optonline.net> writes:

> On 08/25/2017 08:31 PM, Stephen Grant Brown wrote:
>
>  Hi All,
>  How do I setup xsane to use the tesseract OCR engine?
>  I see gocr under preferences->setup->ocr.
>  Yours Sincerely
>  Stephen Grant Brown.
>
> Unless it has been vastly improved, you might as well copy the document by hand! Finding and fixing all the mistakes is not worth the
> trouble!
> Abbyy for Windows does an excellent job. One of only two programs I will boot Windows for. (The other one is a phono-to-CD program.)

My experience OCRing a 16 page document with tesseract last spring was
quite good.  I didn't try to set xsane up to do it (as I thought it
would be a *long* time before I did it again), I scanned the document to
ppm files, sent them to tesseract, put the output of tesseract into a
.txt file, and cleaned up from there.  While it wasn't perfect, it was
far better than retyping the whole thing would have been.
-- 
"Erwin, have you seen the cat?" -- Mrs. Shrödinger

Back to linux.debian.user | Previous | Next — Previous in thread | Next in thread | Find similar | Unroll thread


Thread

xsane & tesseract "Stephen Grant Brown" <steve.brown_nbn@iinet.net.au> - 2017-08-26 03:50 +0200
  Re: xsane & tesseract Doug <dmcgarrett@optonline.net> - 2017-08-26 04:30 +0200
    Re: xsane & tesseract Joe Pfeiffer <pfeiffer@cs.nmsu.edu> - 2017-08-26 06:30 +0200
      Re: xsane & tesseract Siard <shiems146@kpnplanet.nl> - 2017-08-26 10:50 +0200

csiph-web