by yevanchen · v0.0.2
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
A powerful PDF text extraction plugin for Dify powered by PyMuPDF (aka fitz).
PyMuPDF Plugin is a high-performance tool that allows you to extract, analyze, and manipulate text content from PDF documents directly within Dify applications. Built on the robust PyMuPDF library, this plugin provides accurate and efficient PDF text extraction capabilities.
To install the PyMuPDF Plugin:
Once installed, the plugin can be accessed through the Dify interface:
The plugin returns data in multiple formats:
{
"example.pdf": [
{
"text": "Content from page 1...",
"metadata": {
"page": 1,
"file_name": "example.pdf"
}
},
{
"text": "Content from page 2...",
"metadata": {
"page": 2,
"file_name": "example.pdf"
}
}
]
}
This plugin does not collect, store, or transmit any user data beyond what is necessary for processing the provided PDF files. All processing is done within the plugin execution environment, and no data is retained after processing completes.
This plugin is licensed under the AGPL-3.0 License.
For questions, support, or feedback, please contact:
This plugin is provided "as is" without warranty of any kind, express or implied. Users should ensure they have appropriate rights to process any PDF documents uploaded for extraction.