How Bengali OCR works
Drop a photo of a Bangla newspaper, a scanned book page, an exam paper or a screenshot, and the text is recognised and saved as Unicode Bengali, ready for Word, Google Docs, Facebook or a translator. Because the output is standard Unicode, it does not depend on Bijoy or any other legacy keyboard encoding, and it displays the same on every phone and computer.
Bengali script hangs its letters from a headline, like Devanagari, and builds many consonant clusters (juktakkhor) such as ক্ষ, ঙ্গ and ন্ত, often drawn as compact fused shapes. Vowel signs such as ি and ে are written before the consonant they follow in speech, and marks such as hasanta (্), khanda ta (ৎ) and chandrabindu (ঁ) are small. Tesseract's Bengali model reads whole lines with a neural network trained on printed Bengali, which copes with these shapes far better than OCR that matches one letter at a time.
Juktakkhor printed in older or decorative fonts, where two or three letters are fused into an unfamiliar shape, are the most common source of errors, followed by the small marks in low-resolution images. Bengali digits (০ to ৯) are read as Bengali digits. The Bengali model does not know Latin letters, so for pages that mix in English words, tick Also read English words; for a mostly English page, choose English instead.
The Bengali model, about 1.4 MB, downloads from the jsDelivr CDN the first time you press the button and is cached, along with the OCR engine of about 4 MB, so later images start straight away. The image itself is read inside this tab and never uploaded. The result is a UTF-8 text file named after the image. For an editable Word document instead, use Image to Word and choose Bengali.
Bengali OCR options explained
| Language | Bengali, already selected. |
|---|---|
| Script | Bengali (Bangla), written left to right, with a headline joining the letters. |
| Model size | About 1.4 MB, downloaded once and cached. |
| Output | Plain UTF-8 text, one file per image. |
| Layout | Not kept; the words come out in reading order. |
When to use it
Use Bengali OCR to copy text from a scanned book or newspaper, to type up a photographed notice, circular or application, to search a collection of old documents, or to paste Bangla text from an image into a translator.
Use clear print at a good size, photographed straight on. Conjuncts need resolution: text at least 20 pixels tall is recognised much better than tiny print.
About the OCR engine
Text recognition uses Tesseract, the open-source OCR engine first developed at HP, then for many years at Google, and now maintained by its open-source community. Its current generation reads whole lines of text with a neural network (an LSTM) rather than matching letters one at a time, which is what lets it handle joined scripts such as Arabic and Devanagari. Tesseract.js compiles it to WebAssembly so it runs in this tab. Each language has its own trained model, downloaded from the jsDelivr CDN only when that language is chosen and then cached, so reading English never downloads Hindi, and the other way round. The models used here are the integer versions of Tesseract's most accurate models, which keep nearly all of their accuracy at a fraction of the size. Tesseract is built for printed text: it reads books, letters, forms, signs and screenshots well, handwriting poorly, and it keeps the words of a page in reading order but not its layout.
Bengali OCR troubleshooting
Conjuncts come out as separate letters
The print is too small or blurred for the joined letters to be recognised. Use a sharper, larger image.
The result is English gibberish
English was selected. Choose Bengali and convert again.
A Bijoy PDF copies as broken text
PDFs typed in legacy Bijoy fonts do not copy as Unicode. Convert the page to an image and read it here to get proper Unicode Bangla.
How to use Bengali OCR, step by step
- Press "Choose images" or drag files onto the box.
- Set the options if you need to; the defaults suit most uses.
- Press "Extract text". The work happens on your device.
- Save the result with its download button.
Is it safe to do this online?
With most online tools, "online" means your file is uploaded to a company's server, processed there and kept for a while before it is deleted. Here it is not. The page downloads the tool's code to your browser, and your file is read and processed inside the tab on your own device. It is never sent to TapToConvert or anyone else.
You can check this yourself: once the page and its engine have loaded, turn off Wi-Fi and the tool still works. That also means there is no queue, no daily limit and no file size cap set by a server; the only limit is the memory your browser gives a single tab.
Bengali OCR FAQ
Is the output Unicode or Bijoy?
Unicode, which works everywhere without special fonts.
Can it read handwritten Bangla?
Not reliably; it is trained on printed text.
Does it work for Assamese?
Assamese shares the script but has its own letters and model. Choose Assamese in the language list.
Can I read several pages at once?
Yes. Add all the images; each gets its own text file.
Is it free?
Yes. No sign-up, no watermark and no limit on use. TapToConvert is supported by advertising.
Are my files uploaded?
No. Everything happens inside your browser tab, on your own device.