how conversion works

Structure first. Fidelity honestly.

anydoctomd uses format-aware browser parsers instead of treating every file as plain text. The output is designed to be useful and inspectable, not to pretend that a visual document and Markdown are identical.

the short version

Four stages, one reviewable result.

1 · identify

Detect the source format and choose the format-aware parser.

2 · structure

Recover headings, paragraphs, lists, tables, page sections, and links where available.

3 · normalize

Clean the result into readable Markdown without pretending to preserve every pixel.

4 · review

Show the output for editing, copying, downloading, or a later API workflow.

free browser workspace

Local by default for Convert.

Supported conversion happens in the current tab. Browser memory, device speed, available assets, and format support define the practical ceiling.

developer option

Server-side when you need it.

The API is useful for keys, batch runs, persistent sources and artifacts, and integrations. It is a separate path with different storage and trust considerations.

Office and OpenDocument files

Modern DOCX, PPTX, and XLSX files are ZIP packages containing XML parts. The converter reads document paragraphs, styles, relationships, slide sections, worksheets, shared strings, formulas, and table cells. ODT, ODS, and ODP use their XML content and format-specific structures.

PDF and OCR

Text PDFs are read through their text layer and grouped into page sections using coordinates. When a PDF has no extractable text, the browser can rasterize pages and use OCR. Image OCR follows the same browser-first approach.

HTML, EPUB, CSV, RTF, and text

HTML extraction removes common interactive chrome before collecting headings, paragraphs, lists, quotes, code, and tables. EPUB chapters follow the package reading order. CSV parsing respects quoted fields and common delimiters. RTF and plain text use lightweight state-aware normalization.

review before relying on it

Known limitations

  • • Visual positioning, animations, floating objects, charts, and embedded media are not fully reproduced.
  • • OCR can misread low-resolution scans, handwriting, columns, names, and numbers.
  • • URL conversion depends on public access, CORS, and the HTML returned by the target site.
  • • AI-ready means cleaned and structured; it does not mean summarized, fact-checked, or error-free.
  • • Large files and complex layouts can be limited by browser memory or parser support.

references

Technical sources.