Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > linux.debian.user > #185963
| From | Joe Pfeiffer <pfeiffer@cs.nmsu.edu> |
|---|---|
| Newsgroups | linux.debian.user |
| Subject | Re: xsane & tesseract |
| Date | 2017-08-26 06:30 +0200 |
| Message-ID | <uiATn-4Zy-1@gated-at.bofh.it> (permalink) |
| References | <uiyox-3f2-3@gated-at.bofh.it> <uiz1f-3Ni-3@gated-at.bofh.it> |
| Organization | A noiseless patient Spider |
Doug <dmcgarrett@optonline.net> writes: > On 08/25/2017 08:31 PM, Stephen Grant Brown wrote: > > Hi All, > How do I setup xsane to use the tesseract OCR engine? > I see gocr under preferences->setup->ocr. > Yours Sincerely > Stephen Grant Brown. > > Unless it has been vastly improved, you might as well copy the document by hand! Finding and fixing all the mistakes is not worth the > trouble! > Abbyy for Windows does an excellent job. One of only two programs I will boot Windows for. (The other one is a phono-to-CD program.) My experience OCRing a 16 page document with tesseract last spring was quite good. I didn't try to set xsane up to do it (as I thought it would be a *long* time before I did it again), I scanned the document to ppm files, sent them to tesseract, put the output of tesseract into a .txt file, and cleaned up from there. While it wasn't perfect, it was far better than retyping the whole thing would have been. -- "Erwin, have you seen the cat?" -- Mrs. Shrödinger
Back to linux.debian.user | Previous | Next — Previous in thread | Next in thread | Find similar | Unroll thread
xsane & tesseract "Stephen Grant Brown" <steve.brown_nbn@iinet.net.au> - 2017-08-26 03:50 +0200
Re: xsane & tesseract Doug <dmcgarrett@optonline.net> - 2017-08-26 04:30 +0200
Re: xsane & tesseract Joe Pfeiffer <pfeiffer@cs.nmsu.edu> - 2017-08-26 06:30 +0200
Re: xsane & tesseract Siard <shiems146@kpnplanet.nl> - 2017-08-26 10:50 +0200
csiph-web