Skip to main content

Free OCR tool

Image to Word Converter

Extract the text from a picture and download it as a .docx you can open in Word, Google Docs, LibreOffice, or Pages. What you get is the words as clean editable paragraphs — not a pixel-for-pixel recreation of the original page — and knowing that difference up front saves a lot of frustration.

Must match the text in your image (e.g. English for Latin letters).

PNG, JPG, WEBP, GIF · up to 10 images · max ~8 MB each

Extracted text will appear here.

Illustration of a photographed page becoming an editable Word document
Illustration of a photographed page becoming an editable Word document

Walkthrough

Reusing a page from a printed report

A colleague sends a photo of two pages from a printed report and asks you to work the text into a document you are drafting.

  1. Photograph or crop each page separately so one page fills the frame, rather than shooting the open spread as one wide image.
  2. Extract the text, then read it through in the results box and fix anything obviously wrong before you download.
  3. Choose the DOCX download to get an .docx file, and open it in Word or Google Docs.
  4. Apply your own heading and body styles to the paragraphs, then re-create any table or figure by hand from the original image.

What to expect: Expect the body prose to arrive in good shape and need little more than a read-through. Expect to spend your time on structure rather than words: headings arrive as ordinary paragraphs, and anything that was laid out in columns or a table will need rebuilding.

Takeaway: OCR saves you the typing, not the formatting. Budget your effort for restyling rather than retyping, and the workflow makes sense.

What the Word file actually contains

It helps to know exactly what is inside the file you download, because it explains everything about how the workflow feels. The .docx is a real Office Open XML package containing one document part, and that part holds your text as a plain run of paragraphs. Every line in the results box becomes a paragraph. Empty lines become empty paragraphs, so the vertical spacing of the original is roughly preserved. The page is set to US Letter with one-inch margins.

What is deliberately absent is any character or paragraph styling. There is no bold, no italic, no font size variation, no colour, no indentation, no lists, no tables. Your word processor renders the whole thing in its default body style, which for most installations means Calibri or Aptos at eleven points.

That sounds like a limitation, and in one sense it is, but it is also what makes the file pleasant to work with. A document with no formatting is a document with no formatting to undo. You select all, apply your own style set, and you are working with clean material rather than fighting someone else's leftover markup.

Why layout reconstruction is a trap

The obvious question is why a converter would not try to rebuild the original appearance, given that recognition already knows where every word sat on the page. The answer is that knowing coordinates is not the same as knowing intent.

Consider a heading. Recognition can tell you a line of text is larger and bolder than the surrounding text, but deciding whether that makes it a Heading 1, a Heading 2, or a bold lead-in sentence requires understanding the document's structure. Get it wrong and you have a document whose outline is subtly incorrect, which is more annoying than one with no outline at all.

Tables are worse. Visually, a table is text aligned into columns, sometimes with ruled lines and sometimes with nothing but whitespace. Reconstructing one means deciding where the cell boundaries are, and an engine that guesses wrong produces a table with merged and split cells in the wrong places. Repairing that in Word takes longer than building the table fresh from the original image in front of you.

So the design choice here is to give you the text reliably and let you own the structure. If a page is mostly prose, you lose very little. If a page is mostly table, you were always going to be doing manual work.

Getting from plain paragraphs to a finished document

The fastest route is to style top-down rather than fixing line by line. Open the .docx, select everything, and apply your body text style first so the whole document has a consistent base. Then walk through and promote the lines that should be headings, which is quick because they are usually short lines sitting alone between blank ones.

Lists need a small amount of attention. Bullet characters and numbers from the original page come through as literal text at the start of the line — an actual bullet glyph or the characters "1." — rather than as Word list formatting. Selecting those paragraphs and applying a list style, then deleting the leftover characters, converts them properly. Word's find-and-replace with wildcards can strip leading numbering across a long document in one pass.

Keep the original image open beside the document while you do this. You need it anyway to rebuild tables and figures, and having it visible makes it natural to spot-check words as you go, which is the proofreading step that OCR always requires.

Accuracy still governs the result

The Word file is only ever as good as the recognition behind it, and the export step cannot improve a poor extraction. Everything that determines OCR quality applies before you ever click download: how sharp the photograph is, how evenly lit the page is, whether the paper was flat, and how much of the frame the text occupies.

The single most useful habit is to read the extracted text in the results box before downloading, not after. Correcting a word there takes a second. Correcting it after you have styled a fifteen-page document means finding it again amongst your own edits.

Pay particular attention to anything the surrounding language cannot vouch for. Ordinary prose is largely self-correcting, because a model that misreads a letter in a common word will usually still produce the right word. Names, reference codes, part numbers, and figures have no such safety net, and they are exactly the content people most regret not checking.

When Word is the right destination

This route earns its place whenever the text is heading somewhere it will be edited by a person. Quoting a passage from a printed source into an essay, lifting a section of a report into a new draft, turning a photographed letter into something you can amend and resend, or getting a printed form's wording into a document you can revise — all of these are jobs where a .docx is the natural next step.

It is the wrong choice when the destination is a machine. Pulling figures into a spreadsheet, feeding text into a script, or pasting into a note-taking app all work better from plain text, where you are not carrying document markup that the receiving application will only strip out again.

And it is the wrong choice when what you actually need is the page itself rather than its words. Archiving a document, keeping a receipt as evidence, or sending someone a copy of an original all call for the image or a scan, because those preserve the thing rather than a transcription of it.

Where the file is created

The .docx is assembled in your browser from the text sitting in the results box. Nothing is uploaded in order to build it, and the file does not pass through our server on its way to your downloads folder.

The recognition step before it is the part that involves a network request. By default the image is sent through our server to Google's Gemini API, which is what makes the extraction accurate on difficult inputs. Turning on Local OCR only before you extract keeps the image in your browser and uses Tesseract.js instead, at some cost in accuracy on faint or unusual text. Either way, the resulting Word file is produced locally.

The Privacy Policy and Data Retention Policy set out exactly what happens on each path.

Where to go next

If you want the same text as a plain document instead, Image to Text covers the general workflow, and Image to PDF explains what the PDF download does and does not produce. For multi-page printed material that already exists as a PDF, start at PDF to Text. If your source is a photograph that came out poorly, the guide to improving OCR on low-quality images is the place to look before trying again.

Frequently asked questions

Will the Word file look like the original page?

No, and it is not trying to. The .docx contains your extracted text as a sequence of ordinary paragraphs in the default font, with blank lines preserved where they appeared. Fonts, sizes, bold and italic, colours, columns, tables, headers, footers, and images from the original are not reproduced. If you need a visually faithful copy of a page, what you want is a scan or photo of it, not an OCR conversion — those are genuinely different jobs.

Why not preserve the formatting? Other converters claim to.

Some paid desktop products do attempt layout reconstruction, and on clean, simple, well-scanned pages they can do a reasonable job. The reason we do not is that partial layout recovery tends to be worse than none: you end up with a document full of stray text boxes, odd tab stops, and hard line breaks that take longer to clean up than starting from plain paragraphs. Clean text you style yourself is faster to work with, and it is honest about what was actually recovered.

Does the .docx open properly in Google Docs and LibreOffice?

Yes. The file is a standard Office Open XML document, which is the same format Word itself writes, so Google Docs, LibreOffice Writer, Pages, and mobile Office apps all open it directly. It is generated in your browser rather than on a server, so no upload step is involved in producing the file.

Does it handle Hindi, Tamil, or other non-Latin scripts?

Yes. The .docx is UTF-8 throughout, so any script the recognition step managed to read is stored correctly in the file and will display properly as long as your word processor has a font for it. This is worth knowing because the PDF download is more limited in this respect — for non-Latin text, DOCX and TXT are the reliable choices.

Should I choose DOCX or plain TXT?

Choose DOCX when the text is going into a document you will edit and style — a report, an essay, a letter. Choose TXT when the destination is code, a spreadsheet, a note-taking app, or anything that will re-parse the content itself, since TXT avoids a layer of markup you do not need. The words are identical either way; only the container differs.

Can I convert several images into one Word document?

Queue your images, extract them, and the results accumulate in the text box in the order they were processed. One DOCX download then gives you everything in a single file. Ten images per batch is the limit, and the daily AI OCR allowance is ten runs per day per IP address, after which browser-based recognition continues to work.

Ready to extract text?

Use the tool at the top of this page — free, no signup, AI OCR with browser fallback.

Back to the OCR tool