by fangyong20062006 · v0.0.2
Convert digital-native PDFs into clean Markdown or structured JSON locally with PyMuPDF — detects headings, tables, and reading order, strips repeated headers/footers, and supports page-range selection. Pure Python, no model files, runs on amd64 and arm64.
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
Convert digital-native PDFs into clean Markdown or structured JSON locally, using PyMuPDF.
Add the tool to your Dify workflow and provide a PDF file. The plugin reads the uploaded file and returns clean Markdown or structured JSON.
This plugin extracts the text layer of digital-native PDFs and produces good table results. It does not perform OCR, so scanned/image-only PDFs are not supported.
The plugin downloads the file you provide (e.g. a file uploaded to your Dify instance) into memory and parses it locally with PyMuPDF. No file content is sent to any external service. See PRIVACY.md for details.