logo
Welcome Guest! To enable all features please Login or Register.

Notification

Icon
Error

New Topic Post Reply
Options
Go to last post Go to first unread
Akshara  
#1 Posted : Monday, April 3, 2017 12:04:30 AM(UTC)
Quote
Akshara

Rank: Member

Groups: Registered
Joined: 2/7/2017(UTC)
Posts: 12
India
Location: Pune

Hey Paul,


I have a query about OCR on PDF where the PDF used for OCR is not in English. Is there any solution regarding doing OCR on PDF of different Languages.

Waiting for your response.

Thanks & Regards
Akshara
Paul Rayman  
#2 Posted : Monday, April 3, 2017 8:58:53 AM(UTC)
Quote
Paul Rayman

Rank: Administration

Groups: Administrators
Joined: 1/5/2016(UTC)
Posts: 1,138

Thanks: 10 times
Was thanked: 133 time(s) in 130 post(s)
Please see this overload method
https://tesseract.pataga...es_Ocr_OcrApi_Init_1.htm

or remarks section here
https://tesseract.pataga...es_Ocr_OcrApi_Init_2.htm


You have two way:
1.
Languages[] langs = { Languages.English, Languages.Swedish};
api.Init(langs);


2.
api.Init(null as string, "eng+swe");
Please note - first parameter it is a path to tessdata folder

Also do not forget to download the apropriate languages at the link below and to add them into tessdata folder
https://tesseract.patagames.com/langs/
Quick Reply Show Quick Reply
Users browsing this topic
Guest (4)
New Topic Post Reply
Forum Jump  
You can post new topics in this forum.
You can reply to topics in this forum.
You can delete your posts in this forum.
You can edit your posts in this forum.
You cannot create polls in this forum.
You can vote in polls in this forum.