Processing Files

Converts PDF files to markdown text. Use when the user wants to extract text from PDFs, convert PDFs to readable format, or process PDF documents.

tomevault-io Updated

File contents

File Processing Tools

Tools for converting and processing file formats.

PDF to Markdown

Extract text from PDF files and convert to markdown format.

Basic usage

result = await pdf_to_markdown(file_path="/path/to/document.pdf")
# Returns: {"success": True, "data": {"markdown": "# Document Title\n\nContent..."}}

With page limits

# Extract first 5 pages only
result = await pdf_to_markdown(
    file_path="/path/to/document.pdf",
    max_pages=5
)

Response format

{
  "success": true,
  "data": {
    "markdown": "extracted markdown content",
    "page_count": 10,
    "file_path": "/path/to/document.pdf"
  }
}

Requirements

Requires pymupdf package for PDF processing.

When to use

  • Converting PDF reports to readable text
  • Extracting content from PDF documents
  • Processing PDF files for further analysis

Converted and distributed by TomeVault — claim your Tome and manage your conversions.

tomevault-io/skills-registry/tree/main/binome-dev--humcp--files commit abad63b966

Frequently asked questions

npx skillmds@latest add tomevault-io/processing-files