Home › Search & Research › pdfmux
pdfmux
PDF-to-Markdown extraction that audits its own output and flags any extractor's silent drops.
Topics: Search & Data Extraction
Connect
Review any command before running it. Package names and URLs come from the server's own registry entry.
Package (pypi 1.8.7)
uvx pdfmux
Or add to your MCP client config:
{
"mcpServers": {
"pdfmux": {
"command": "uvx",
"args": [
"pdfmux"
],
"env": {
"GEMINI_API_KEY": "<YOUR_VALUE>"
}
}
}
}
GEMINI_API_KEYsecret — Optional Google AI API key. Enables the paid Gemini Flash fallback for pages where all local backends report low confidence. pdfmux runs 100% locally without this.
This server takes extra arguments — see its repository.
Related servers
Firecrawl MCP Server
MCP server for Firecrawl — web search, scraping, and biomedical/arXiv paper search.
wigolo
Local-first web intelligence MCP server for AI coding agents
exa-mcp-server
Connect AI agents to Exa for web search, content fetching, and multi-step research.
arxiv-mcp-server
Search arXiv papers, download full text, semantic search, citation graphs, and alerts via MCP.
trieve
Crawl, embed, chunk, search, and retrieve information from datasets through [Trieve](https://trieve.ai)
brightdata-mcp
Bright Data's Web MCP server enabling AI agents to search, extract & navigate the web
Listed in punkpeye/awesome-mcp-servers (MIT)
Data from the Official MCP Registry