Sanchaya : A great project with clear vision!
Now about my first experience of its OCR. I tried it on a text containing Sanskrit, Hindi and English. Accuracy seems as good as it has been for Tesseract for these languages. Speed is also moderate. I think processing is being done on the client side.
The menu is easy to grasp. The options are quite many !
Some improvements i would like to suggest :
(1) The find-replace option could also have option for regular expressions. This is important because OCRed text has good chances of having systematic errors instead of random errors.
(2) some way to tell about selecting a table and telling about the number of columns. (I did not try on pages with two or three columns of text)
(3) It left two initial words unboxed. There could be some way to force it to 'see' those texts.
(4) 'train data' has no response. Is it active in the present version?
(5) It shows 'Page 10 of 24' at the top of left hand panel. But in reality, it displays two pages (page 10 & 11) and produces text for these displayed pages if 'recognise' is clicked. Instead of 'recognise' it could be 'recognise selected' or something like that.
-- अनुनाद