by fangyong20062006 · v0.0.2
Recursively crawl web pages from a starting URL and depth, returning structured JSON for each page (title, main content as Markdown, links, and metadata). Supports same-domain restriction, max-page limits, and URL include patterns; tolerant of self-signed intranet certificates.
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
Recursively crawl web pages from a starting URL for use in Dify workflows.
Add the Web Crawler tool to your Dify workflow, provide a starting URL and crawl depth, and receive structured page data.
This plugin sends HTTP requests to the URLs you provide and fetches their content. The fetched web pages are processed within your Dify runtime and are not sent to any third-party service by the plugin. You are responsible for ensuring you have permission to crawl the target sites. See PRIVACY.md for details.