crawl_deep
Crawl websites deeply using breadth-first search
map_site
Discover and map website structure
extract_content
Extract and analyze main content from web pages with enhanced readability detection
process_document
Process documents from multiple sources and formats including PDFs and web pages
summarize_content
Generate intelligent summaries of text content with configurable options
analyze_content
Perform comprehensive content analysis including language detection and topic extraction
extract_structured
Extract structured data from a webpage using LLM-powered analysis and a JSON Schema. Falls back to CSS selector extraction when no LLM provider is configured.
batch_scrape
Process multiple URLs simultaneously with support for async job management and webhook notifications
scrape_with_actions
Execute browser action chains before scraping content, with form auto-fill and intermediate state capture
deep_research
Conduct comprehensive multi-stage research with intelligent query expansion, source verification, and conflict detection
track_changes
Enhanced content change tracking with baseline capture, comparison, scheduled monitoring, advanced comparison engine, alert system, and historical analysis
generate_llms_txt
Analyze websites and generate standard-compliant LLMs.txt and LLMs-full.txt files defining AI model interaction guidelines
stealth_mode
Advanced anti-detection browser management with stealth features, fingerprint randomization, and human behavior simulation
localization
Multi-language and geo-location management with country-specific settings, browser locale emulation, timezone spoofing, and geo-blocked content handling