The Word reader, the Markdown and HTML converters, the PDF writer and the PDF text extractor all run inside this page, so your files and inputs never leave your device. The size a document can reach is limited by the memory your tab has, not by an upload cap.
Related tools
- PDF ToolboxMerge, split, rotate, reorder, watermark, fill and sign PDFs in your browser.
- Document ScannerTurn a phone photo of a page into a straightened, cropped, multi-page PDF.
- Font SubsetterStrip a font down to the characters you use and write it back out as WOFF2, WOFF, or OpenType.
- HTML to MarkdownTurn HTML and rich text pasted from Google Docs or Word into tidy Markdown.
- Lorem Ipsum GeneratorGenerate placeholder paragraphs, sentences, or words in classic Latin or plain English filler, as plain text, HTML, or Markdown.
- OCRPull text out of images with tesseract, running in your browser.
What it does
Converts documents between the four formats people actually pass around: Word .docx files, Markdown, HTML, and plain text, plus a PDF on the way out. DOCX comes in through a real Open XML reader, so headings, lists, tables, bold and italic survive, and pictures come across as inline images in the HTML output. The PDF path is a clean text flow rendering built from the document structure: headings sized by level, bullet and numbered lists, indented quotes, monospaced code blocks, horizontal rules, tables flattened to rows, word wrapping measured against the real font, page breaks and page numbers. Dropping a PDF in runs the other direction and extracts its text, tidied up so hyphenated line breaks and column padding do not follow you out.
How to use it
Drop a .docx, .md, .html, .txt or .pdf file onto the panel, or paste Markdown or HTML straight in. The input format is detected for you, and you can override it if the guess is wrong. Pick an output format, check the preview pane, then copy the result or download it. For PDF output you can set A4 or US Letter, body text size, page margin, and whether pages get numbered.
Why this one
The popular document converters upload your file to a server, queue it, and hand it back with a watermark or a two file daily cap unless you pay. That is a rough deal for the documents people convert most: contracts, resumes, invoices, medical letters, drafts nobody else should read. This one does the whole conversion in the page, so your files and inputs never leave your device, with no account, no queue and no size cap beyond what your own browser can hold. It is also honest about the one thing it does differently: the PDF is a text flow rendering rather than a browser screenshot, and the Print to PDF button is right there when you need the exact page.
FAQ
- Is my document uploaded anywhere?
- No. The Word reader, the Markdown and HTML converters, the PDF writer and the PDF text extractor all run inside this page, so your files and inputs never leave your device. The page keeps working after the first load even with the network off.
- Why does the PDF look plainer than my original?
- Because it is a text flow rendering, not a screenshot of a styled page. The converter reads the document structure (headings, lists, quotes, code, tables) and lays it out with proper wrapping, page breaks and page numbers, but it does not run a CSS engine, so colors, columns, background art and web fonts are not reproduced. When you need the page exactly as it looks on screen, use the Print to PDF button, which hands the preview to your browser's own print dialog.
- Does it keep images from my Word file?
- In the HTML output, yes: pictures embedded in a .docx come across as inline images, so the HTML is self contained. In the PDF output, no: the text flow renderer skips images and leaves a short note where each one sat, so you can see what was dropped. Markdown output keeps the image reference, and plain text output carries the same short note.
Keyboard shortcuts: press ? anywhere on this page to see them.