Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I’ve found that none of the open source stuff works well for Japanese language documents. Most of the time, I’ve just ran them through Adobe Acrobat’s OCR and dumped the results into a text file. There are still mistakes, but it at least returns a passable result compared to others.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: