by us · v0.0.1
CRW — the open-source web scraper built for AI agents. Scrape, crawl, and map websites. Firecrawl-compatible API, 5.5x faster, 75x less memory. Self-hosted or fastcrw.com cloud.
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
Web scraping, crawling, and URL mapping for Dify workflows and agents.
CRW is an open-source web scraper built for AI agents. Firecrawl-compatible API, 5.5x faster, 75x less memory. Works with fastcrw.com cloud or any self-hosted CRW instance.
| Tool | Endpoint | Description |
|---|---|---|
| Scrape | POST /v1/scrape | Scrape a single URL and return clean markdown, HTML, plain text, or structured JSON |
| Crawl | POST /v1/crawl | Start an async BFS crawl with depth and page limits |
| Crawl Status | GET /v1/crawl/{id} | Check crawl job status or cancel a running job |
| Map | POST /v1/map | Discover all URLs on a website via link extraction and sitemap parsing |
Install from the Dify Marketplace, or load locally during development:
# Package the plugin
dify plugin package ./crw
# Install the .difypkg file via Dify Settings > Plugins
Sign up at fastcrw.com and get 500 free credits:
crw_live_... from fastcrw.comcurl -fsSL https://raw.githubusercontent.com/us/crw/main/install.sh | bash
crw # starts on http://localhost:3000
any-value (or leave empty if no auth)http://localhost:3000docker run -d -p 3000:3000 ghcr.io/us/crw:latest
Same as Option B.
Extract clean content from a single URL. Supports:
Start an async breadth-first crawl from a URL:
Check on or cancel a running crawl job:
Discover all URLs on a website: