Skip to content
CUSATPublic
forked from HairyFotr/OCRTrain

About

Overfit OCR to your dataset with genetic algorithms.

Resources

Stars

0 stars

Watchers

1 watching

Forks

 
 

Latest commit

 

History

5 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

-------------
 OCR Trainer
-------------

1. Put your images into img/
2. Put your transcriptions into a text file... "filename transcription"
3. Run runTrain <yourTextFile>
4. Wait and wait and wait :)

Things you need:
  sudo apt-get install imagemagick tesseract-ocr gocr cuneiform ocrad

Things you should probably set (a.k.a. things I should abstract away into files):
  allowedCharacters / stringFilter
  the appropriate string scoring algorithm
  different sequence of param-changing algorithms

About

Overfit OCR to your dataset with genetic algorithms.

Resources

Stars

0 stars

Watchers

1 watching

Forks

Releases

Packages

Contributors