document-parser
MCP serverDocs & knowledgeParse and extract structured data from various document formats (PDF, Word, HTML).
This server has no hosted endpoint yet, so ahel can't serve it. You can still add it. It stays paused until ahel can serve it.
Serve it through your gateway
One link, every agent. Your own credentials, stored once.
Tools it gives your agents (5)
| Tool | What it does |
|---|---|
| parse_pdf | Extract text, tables, and metadata from PDF files with layout preservation. Perfect for agents processing reports, invoices, contracts, research papers. |
| parse_image_text | Perform OCR on images to extract text with confidence scores. Supports screenshots, scanned documents, photos of text. Returns structured text with confidence metrics. |
| html_to_markdown | Convert HTML documents to clean, structured markdown. Preserves headings, links, tables, lists. Perfect for agents that need to process HTML content in LLM-friendly format. |
| extract_tables | Extract tables from any supported document format as structured JSON. Handles PDF tables, HTML tables, CSV-like structures in text. |
| summarize_document | Parse any document and generate a structured summary with configurable detail level. Extracts key information, main points, and metadata. |
Signals
- Last commit
- Apr 2026
- Weekly downloads
- 66
- Tools captured
- 5