DOCX to Text
Turn a .docx into plain text you can search, paste and diff. Paragraphs, line breaks, tabs and table rows are preserved, and headers, footers, footnotes and comments can be pulled in as well. The document is opened and read inside your browser, so it is never uploaded.
Each table row becomes one line, with cells separated by a tab.
Adds a [Header] section for every header part in the document.
Adds a [Footer] section for every footer part in the document.
Comments are listed at the end, each labelled with its author.
How docx to text works
- Add your documentDrop in a .docx, or browse your device. The file is checked to make sure it really is a ZIP-based Office document before anything is parsed.
- Choose what to includeKeep table rows, and switch on headers, footers, footnotes or comments if they matter to what you are extracting.
- Copy or download the textThe result appears in an editable box you can select from, with a plain-text download underneath.
What you get
Pull the words out of a Word document, tables and all. Everything happens inside this page: the file is read by your browser, transformed in memory and handed straight back to you as a download. There is no upload queue, no waiting for a server, and nothing left behind when you close the tab.
Supported formats
Furtu accepts .docx files up to 100 MB each — one file at a time.
- Every file is checked against its real file signature, not just its name, so a renamed or corrupt file is rejected with an explanation instead of failing halfway through.
- Files over the limit are refused up front rather than after a long wait.
Limitations, stated up front
- Text inside text boxes, shapes, charts and SmartArt is not extracted; only the main document flow, tables, and the parts you switch on.
- Images are ignored entirely. A document that is mostly scans produces very little text.
- Tracked deletions are left out and tracked insertions are included, which matches what you would see with changes shown as final.
- Formatting, styles, numbering and hyperlinks are discarded — this is a text extraction, not a converter.
Why this is worth doing locally
The files people most want converted are the ones they least want uploaded: employment contracts, medical letters, bank statements, a dissertation chapter. Reading a ZIP container is also cheap enough to do properly in a browser tab, so there is no reason to accept an upload form for it.
Frequently asked questions
Are my Word documents uploaded anywhere?
No. A .docx is a ZIP archive, and Furtu opens that archive and reads the document XML inside it using code that is compiled into this page. Nothing leaves your device, and nothing is kept after you close or reload the tab.
What happens to formatting?
It does not survive, because plain text has no formatting. What does survive is structure: paragraph breaks become newlines, tabs stay tabs, and each table row becomes a single line with its cells separated by tabs, which pastes cleanly into a spreadsheet.
Why is the text missing from my document?
Almost always because the text is an image. A document made from scans or screenshots has no text layer, and no amount of parsing will produce words from pixels. Furtu says so rather than returning an empty file. Text boxes and shapes are also not included, because Word stores them outside the main document flow.
Can I get the headers and footers as well?
Yes — turn on “Include page headers” and “Include page footers”. They are stored in separate parts of the file, so they are opt-in; without the toggles you get the document body only.
Will it open an old .doc file?
No. A .doc from before 2007 is a completely different binary format, not a ZIP archive. Save it as .docx first — Word, LibreOffice and Google Docs will all do that in one step.