Stage Category: APPLY (Enriches documents with scraped content)Transformation: N documents → N documents (with web content added)
When to Use
When NOT to Use
Parameters
Configuration Examples
Output Schema
Markdown Output (default)
Full Extraction
Error Case
Firecrawl Features
Performance
Common Pipeline Patterns
Enrich Documents with Referenced Content
Scrape and Summarize
Error Handling
Rate Limits and Best Practices
- Batch wisely: Limit to 5-10 URLs per pipeline run
- Cache results: Consider storing scraped content
- Respect robots.txt: Firecrawl handles this automatically
- Use timeouts: Set appropriate
timeout_msfor your use case
Related
- External Web Search - Search the web (Exa)
- API Call - General HTTP enrichment
- Document Enrich - Collection joins

