document-parser

MCP serverDocs & knowledge

Parse and extract structured data from various document formats (PDF, Word, HTML).

This server has no hosted endpoint yet, so ahel can't serve it. You can still add it. It stays paused until ahel can serve it.

Serve it through your gateway

One link, every agent. Your own credentials, stored once.

Tools it gives your agents (5)

ToolWhat it does
parse_pdfExtract text, tables, and metadata from PDF files with layout preservation. Perfect for agents processing reports, invoices, contracts, research papers.
parse_image_textPerform OCR on images to extract text with confidence scores. Supports screenshots, scanned documents, photos of text. Returns structured text with confidence metrics.
html_to_markdownConvert HTML documents to clean, structured markdown. Preserves headings, links, tables, lists. Perfect for agents that need to process HTML content in LLM-friendly format.
extract_tablesExtract tables from any supported document format as structured JSON. Handles PDF tables, HTML tables, CSV-like structures in text.
summarize_documentParse any document and generate a structured summary with configurable detail level. Extracts key information, main points, and metadata.

Signals

Last commit
Apr 2026
Weekly downloads
66
Tools captured
5