OCR is often explained as a simple process:
Image → Text
But that's only one possible destination.
A document image might actually need to become:
Image → DOCX
Image → XLSX
Image → JSON
Image → CSV
Image → SQL
Image → Markdown
Image → LaTeX
Image → Mermaid
That's why I find multi-format OCR more interesting than plain text extraction.












