TapToConvert All conversions

Home / Tools / Hebrew OCR

Hebrew OCR

Extract Hebrew text from photos, scans and screenshots with OCR that writes it right to left in Unicode, final letters included. Paste it anywhere. The image is not uploaded.

Drop files hereChoose images Nothing is uploaded. Any size your device can handle.

Add a file from a URL

The file is fetched by your browser straight from that website, not through our servers. It works when the site allows other websites to download its files.

Settings

Add files to begin.

    How Hebrew OCR works

    Drop a photo of a printed Hebrew page, a document, a sign or a screenshot, and the text is recognised and saved as Unicode Hebrew for Word, Google Docs, email or a translator. Lines are stored right to left in logical order, so the text edits and displays correctly in modern apps, and English words or numbers inside a Hebrew sentence keep their own direction.

    Hebrew has 22 letters, five of which take a different final form at the end of a word: כ and ך, מ and ם, נ and ן, פ and ף, צ and ץ. Several letters are easily confused in small print, such as ב and כ, ד and ר, ה and ח, and ו and ז. Tesseract's Hebrew model reads whole lines with a neural network trained on printed Hebrew, and at about 0.6 MB it is one of the smallest language downloads here.

    Most everyday Hebrew is written without vowel points, and that is what the model reads best. Pointed text (niqqud), as in prayer books, poetry and children's books, often comes out with the points missing or misplaced, and cantillation marks are not read. Rashi script and handwriting are beyond it.

    The Hebrew model, about 0.6 MB, downloads from the jsDelivr CDN the first time you press the button and is cached, along with the OCR engine of about 4 MB, so later images start straight away. It is trained on the Hebrew alphabet, so if the text mixes in English words, tick Also read English words and the English model reads alongside it, a little more slowly. The image itself is read inside this tab and never uploaded. The result is a UTF-8 text file named after the image, stored in logical order so it displays right to left. For an editable Word document instead, use Image to Word and choose Hebrew.

    Hebrew OCR options explained

    LanguageHebrew, already selected.
    ScriptHebrew square script, right to left, usually without vowel points.
    Model sizeAbout 0.6 MB, downloaded once and cached.
    OutputPlain UTF-8 text, one file per image.
    LayoutNot kept; the words come out in reading order.

    When to use it

    Use Hebrew OCR to copy text from a scanned letter or document, to get the wording of a certificate or form into an editor, to search photographed book pages, or to paste Hebrew from an image into a translator.

    For a text with niqqud, read the letters here and add the points afterwards if you need them; for everything else, a sharp, straight photo of the page is all it takes.

    About the OCR engine

    Text recognition uses Tesseract, the open-source OCR engine first developed at HP, then for many years at Google, and now maintained by its open-source community. Its current generation reads whole lines of text with a neural network (an LSTM) rather than matching letters one at a time, which is what lets it handle joined scripts such as Arabic and Devanagari. Tesseract.js compiles it to WebAssembly so it runs in this tab. Each language has its own trained model, downloaded from the jsDelivr CDN only when that language is chosen and then cached, so reading English never downloads Hindi, and the other way round. The models used here are the integer versions of Tesseract's most accurate models, which keep nearly all of their accuracy at a fraction of the size. Tesseract is built for printed text: it reads books, letters, forms, signs and screenshots well, handwriting poorly, and it keeps the words of a page in reading order but not its layout.

    Hebrew OCR troubleshooting

    ד and ר or ה and ח are confused

    Those pairs differ by tiny details. Use a larger, sharper image.

    The vowel points are missing

    Niqqud is often dropped. The letters are usually right; add the points by hand if needed.

    Text looks reversed after pasting

    Set the paragraph direction to right to left in your editor; the text is stored correctly.

    How to use Hebrew OCR, step by step

    1. Press "Choose images" or drag files onto the box.
    2. Set the options if you need to; the defaults suit most uses.
    3. Press "Extract text". The work happens on your device.
    4. Save the result with its download button.

    Is it safe to do this online?

    With most online tools, "online" means your file is uploaded to a company's server, processed there and kept for a while before it is deleted. Here it is not. The page downloads the tool's code to your browser, and your file is read and processed inside the tab on your own device. It is never sent to TapToConvert or anyone else.

    You can check this yourself: once the page and its engine have loaded, turn off Wi-Fi and the tool still works. That also means there is no queue, no daily limit and no file size cap set by a server; the only limit is the memory your browser gives a single tab.

    Hebrew OCR FAQ

    Does it read niqqud?

    Partly. Unpointed text reads best; points are often dropped or misplaced.

    Can it read Yiddish?

    Yiddish has its own model. Choose Yiddish in the language list.

    Is the text right to left?

    Yes, stored as Unicode in logical order.

    Can I get a Word document?

    Yes. Use Image to Word with Hebrew selected; paragraphs are set right to left.

    Is it free?

    Yes. No sign-up, no watermark and no limit on use. TapToConvert is supported by advertising.

    Are my files uploaded?

    No. Everything happens inside your browser tab, on your own device.

    Try MP4 to MP3, HEIC, merge PDF or compress video.