Based on the tif you posted, it appears that the resolution is too low
and the "*" characters are too small to give good results.  The FAQ
documentation suggests you double your resolution at least.  That
might be sufficient.

If not, relative to command line options, the thread on bbtesseract
talks about using "nobatch" in the command line.  I tried using
"nobatch" instead of "batch.nochop" and it seems to make a
difference.  I tried both and compared.  I did not really know how
they work.  Maybe someone else can comment ???

The two "4"s in your image touch, so those need to be handled one way
or another per "Training Tesseract".  Ray Smith had a post in the last
few days about using shorter substrings; that might help.

On Aug 20, 9:06 am, Chris <[email protected]> wrote:
> Using Tesseract 2.04 through command line using the default eng
> training files.
> Attempting to make a box file off the following image.  The first line
> is read fine, but the second is ignored.
>
> Is there anything i can do, any parameters i can pass to adjust how
> the box file is created?
>
> I have the tif posted at the link below, along with a zip of the tif.
>
> http://omni19.tripod.com/recclaim/test.ziphttp://omni19.tripod.com/recclaim/test.tif
--~--~---------~--~----~------------~-------~--~----~
You received this message because you are subscribed to the Google Groups 
"tesseract-ocr" group.
To post to this group, send email to [email protected]
To unsubscribe from this group, send email to 
[email protected]
For more options, visit this group at 
http://groups.google.com/group/tesseract-ocr?hl=en
-~----------~----~----~----~------~----~------~--~---

Reply via email to