TapToConvert All conversions

Home / Tools / Russian OCR

Russian OCR

Extract Russian text from photos, scans and screenshots with OCR trained on Cyrillic print. Get editable Unicode text you can copy, search and translate. Nothing is uploaded.

Drop files hereChoose images Nothing is uploaded. Any size your device can handle.

Add a file from a URL

The file is fetched by your browser straight from that website, not through our servers. It works when the site allows other websites to download its files.

Settings

Add files to begin.

    How Russian OCR works

    Drop a photo of a Russian book page, a document, a screenshot or a sign, and the Cyrillic text in it is recognised and saved as Unicode text that you can paste into Word, Google Docs, a search box or a translator, without needing a Russian keyboard.

    Russian is written in Cyrillic, 33 letters of which several look exactly like Latin ones, such as А, В, Е, К, М, Н, О, Р, С, Т and Х, while being different characters. Reading a Russian page with English selected gives a mix of Latin lookalikes and garbage that cannot be searched or spell-checked; the Russian model outputs proper Cyrillic. It reads whole lines with a neural network trained on printed Russian.

    The letter ё is often printed as е, and the model outputs what is on the page, so do not expect it to add the dots. Italic Cyrillic is harder: the italic т looks like m, и like u and д like g, and a line of italics reads less reliably than upright type. Pre-1918 spelling, with ѣ, і and a hard sign at the end of words, is not what the model is trained on. Latin words such as brand names come out as Cyrillic lookalikes (in our test, Google became Сооде) unless Also read English words is ticked.

    The Russian model, about 2.7 MB, downloads from the jsDelivr CDN the first time you press the button and is cached, along with the OCR engine of about 4 MB, so later images start straight away. The image itself is read inside this tab and never uploaded. The result is a UTF-8 text file named after the image. For an editable Word document instead, use Image to Word and choose Russian.

    Russian OCR options explained

    LanguageRussian, already selected.
    ScriptCyrillic, 33 letters, written left to right.
    Model sizeAbout 2.7 MB, downloaded once and cached.
    OutputPlain UTF-8 text, one file per image.
    LayoutNot kept; the words come out in reading order.

    When to use it

    Use Russian OCR to copy text from a scanned contract, certificate or letter, to get quotes out of photographed books, to search screenshots, or to paste Russian text into a translator when you do not have a Cyrillic keyboard.

    For Russian text with English names or terms in it, tick Also read English words. For a page with long passages in each language, crop the parts apart and read each with its own language.

    About the OCR engine

    Text recognition uses Tesseract, the open-source OCR engine first developed at HP, then for many years at Google, and now maintained by its open-source community. Its current generation reads whole lines of text with a neural network (an LSTM) rather than matching letters one at a time, which is what lets it handle joined scripts such as Arabic and Devanagari. Tesseract.js compiles it to WebAssembly so it runs in this tab. Each language has its own trained model, downloaded from the jsDelivr CDN only when that language is chosen and then cached, so reading English never downloads Hindi, and the other way round. The models used here are the integer versions of Tesseract's most accurate models, which keep nearly all of their accuracy at a fraction of the size. Tesseract is built for printed text: it reads books, letters, forms, signs and screenshots well, handwriting poorly, and it keeps the words of a page in reading order but not its layout.

    Russian OCR troubleshooting

    Cyrillic comes out as Latin gibberish

    English was selected. Choose Russian and convert again.

    Italic text is full of errors

    Italic Cyrillic letters resemble Latin ones. Use a larger image, and proofread long italic passages.

    Old books read badly

    Pre-reform spelling and worn type reduce accuracy. Proofread against the page.

    How to use Russian OCR, step by step

    1. Press "Choose images" or drag files onto the box.
    2. Set the options if you need to; the defaults suit most uses.
    3. Press "Extract text". The work happens on your device.
    4. Save the result with its download button.

    Is it safe to do this online?

    With most online tools, "online" means your file is uploaded to a company's server, processed there and kept for a while before it is deleted. Here it is not. The page downloads the tool's code to your browser, and your file is read and processed inside the tab on your own device. It is never sent to TapToConvert or anyone else.

    You can check this yourself: once the page and its engine have loaded, turn off Wi-Fi and the tool still works. That also means there is no queue, no daily limit and no file size cap set by a server; the only limit is the memory your browser gives a single tab.

    Russian OCR FAQ

    Does it work for Ukrainian or Belarusian?

    Choose Ukrainian in the language list for Ukrainian, which has its own letters such as і, ї and є. Belarusian is not in the list; the Russian model reads most of its letters.

    Will it add the dots to ё?

    No. It outputs the letters as printed.

    Can it read handwritten Russian?

    Not reliably; cursive Cyrillic is very different from print.

    Can I get a Word document?

    Yes. Use Image to Word and choose Russian.

    Is it free?

    Yes. No sign-up, no watermark and no limit on use. TapToConvert is supported by advertising.

    Are my files uploaded?

    No. Everything happens inside your browser tab, on your own device.

    Try MP4 to MP3, HEIC, merge PDF or compress video.