AI product Open source

doc7

doc7 is an open-source document-to-Markdown tool that converts PDFs, Office files, scans, screenshots, charts, formulas, and diagrams into AI-ready Markdown using an OpenAI-compatible multimodal vision model. It renders and understands complete pages rather than relying on character extraction, preserving text, tables, formulas, chart and diagram relationships, image meaning, and visible UI state; the video also describes page-level retry and resume. The model can run locally through LM Studio or Ollama or through a remote endpoint, with no required OCR stack or document-processing service. The same binary provides an interactive CLI, batch processing, model checks, MCP support, a Go SDK, and an asynchronous HTTP service, with installers for macOS, Linux, and Windows.

View repository Mentioned in 1 video ↓

What doc7 is used for

1 use taken from transcripts — each links to the moment in the video.

  • Converts documents to Markdown by rendering pages and sending them to a vision model, preserving charts, diagrams, and visible UI state while supporting page-level retry and resume.

Videos mentioning doc7

1 in the library.