Extract plain text from Word .docx files in your browser: paragraphs, tabs and line breaks preserved. No upload, no signup. No upload — runs 100% in your browser
A Word file is a ZIP archive of XML parts, with the body text living in word/document.xml. This tool unzips your file locally (with fflate, a small open-source decompressor), reads that XML, and joins the text runs into plain lines — one line per paragraph, tabs and hard line breaks preserved. The formatting, styles and images stay behind, which is the point: you get the words.
The old binary .doc format (Word 97–2003) is a completely different file layout and is not supported — re-save it as .docx in Word first (Save As → .docx), which any Word version since 2007 does in seconds. If you drop an old .doc on this page you'll get a clear message instead of garbage.
Feeding documents to AI tools that want plain text, indexing contracts into search, pulling the words out of a resume for an ATS check, cleaning up text that someone pasted into Word and buried under three levels of styling. The text layer that comes out is exactly what was typed — no reflow, no smart-quote mangling, nothing invented.
No. Everything on this page runs inside your browser using JavaScript and WebAssembly. Your file never leaves your device — you can even disconnect from the internet after the page loads and the tool still works.
No — only .docx (the XML-based format used since Word 2007). Save the file as .docx first; the tool tells you clearly if you drop an old .doc on it.
This extracts the document body. Headers/footers/footnotes live in separate XML parts and are not included — for most text-extraction uses that's the wanted behavior.
Images are not text, so no. Table cell text is extracted in document order, but the grid structure is not — plain text has no columns.