Markdown Converter

by intellectronica9b0e00ad1b94No licenseListed Oct 8, 2026Updated Oct 8, 2026

Convert documents and files to Markdown using markitdown. Use when converting PDF, Word (.docx), PowerPoint (.pptx), Excel (.xlsx, .xls), HTML, CSV, JSON, XML, images (with EXIF/OCR), audio (with transcription), ZIP archives, YouTube URLs, or EPubs to Markdown format for LLM processing or text analysis.

Instructions onlyDocuments & Office
AI-generated overview

Converts documents and files such as PDF, Word, PowerPoint, Excel, HTML, CSV, images, audio, ZIP and EPub into Markdown using markitdown.

What it does
This skill instructs an agent to run the markitdown command-line tool through uvx to convert supported file types into Markdown. It covers documents, web and data formats, media with EXIF/OCR or transcription, ZIP archives, YouTube URLs and EPub, and can write output to stdout or a file. Options include extension, MIME type and charset hints, Azure Document Intelligence for difficult PDFs, and third-party plugins. The resulting Markdown preserves headings, tables, lists and links for LLM processing or text analysis.
When to use it
Use it when you need to turn PDFs, Office files, HTML, CSV, JSON, XML, images, audio, ZIP archives, YouTube URLs or EPubs into Markdown text. It is suited to preparing documents for LLM processing or text analysis, and to extracting readable content from complex PDFs with Azure Document Intelligence.
Requirements
Requires the uvx runtime to run markitdown; no installation is needed as dependencies are cached on first run. Optional Azure Document Intelligence requires an endpoint and credentials. Network access is needed for uvx dependency retrieval and YouTube URLs. Ships no scripts; instructions only.

Markdown Converter

Convert files to Markdown using uvx markitdown — no installation required.

Basic Usage

bash
# Convert to stdoutuvx markitdown input.pdf
# Save to fileuvx markitdown input.pdf -o output.mduvx markitdown input.docx > output.md
# From stdincat input.pdf | uvx markitdown

Supported Formats

  • Documents: PDF, Word (.docx), PowerPoint (.pptx), Excel (.xlsx, .xls)
  • Web/Data: HTML, CSV, JSON, XML
  • Media: Images (EXIF + OCR), Audio (EXIF + transcription)
  • Other: ZIP (iterates contents), YouTube URLs, EPub

Options

bash
-o OUTPUT      # Output file-x EXTENSION   # Hint file extension (for stdin)-m MIME_TYPE   # Hint MIME type-c CHARSET     # Hint charset (e.g., UTF-8)-d             # Use Azure Document Intelligence-e ENDPOINT    # Document Intelligence endpoint--use-plugins  # Enable 3rd-party plugins--list-plugins # Show installed plugins

Examples

bash
# Convert Word documentuvx markitdown report.docx -o report.md
# Convert Excel spreadsheetuvx markitdown data.xlsx > data.md
# Convert PowerPoint presentationuvx markitdown slides.pptx -o slides.md
# Convert with file type hint (for stdin)cat document | uvx markitdown -x .pdf > output.md
# Use Azure Document Intelligence for better PDF extractionuvx markitdown scan.pdf -d -e "https://your-resource.cognitiveservices.azure.com/"

Notes

  • Output preserves document structure: headings, tables, lists, links
  • First run caches dependencies; subsequent runs are faster
  • For complex PDFs with poor extraction, use -d with Azure Document Intelligence

Source and attribution

Source:intellectronica/agent-skillsinskills/markdown-converterat commit9b0e00a

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal