unclecode/crawl4ai
Documented errors, page 2 of 2. Back to unclecode/crawl4ai
| Code / Message | Type | Severity | Tags |
|---|---|---|---|
| extraction_strategy must be an instance of… | validation | error | validation, extraction, configuration, type-error |
| Query parameter 'q' is required | http | error | llm, http-400, query-params, validation |
| Request timed out | exception | error | timeout, http-crawler, network |
| Failed to parse schema JSON | exception | error | extraction, schema-generation, llm, json-parse |
| stdin is not a terminal | exception | warning | browser-profiler, tty, interactive, environment |
| Request failed | exception | info | artifacts, not-found, ephemeral-storage |
| File not found | exception | warning | ssrf-protection, egress, url-validation, url-blocked |
| Invalid status: . Must be one of: all, active, completed… | http | error | http-400, validation, monitoring, api |
| LLM returned an empty response | exception | error | extraction, schema-generation, llm, empty-response |
| LLM returned empty script. | exception | error | llm, c4a-script, compile, empty-response, runtime-error |
| psutil not available, cannot clean old browser | exception | error | windows, browser-manager, dependencies, psutil |
| virtual_scroll_config must be VirtualScrollConfig object or… | validation | error | validation, virtual-scroll, configuration, type-error |
| Unknown procedure | exception | error | c4a-script, compile, undefined-reference, procedures, validation |
| C4A Script compilation error | validation | error | c4a-script, javascript, compilation, validation |
| URL must start with 'http://', 'https://', 'file://', or… | validation | error | url-validation, scheme, input-validation |
| Invalid email domain | http | error | auth, email, dns, validation |
| webhook header not allowed | validation | warning | crawl4ai, webhook, validation, http-headers, security |
| Failed to connect | exception | error | docker-client, network, connection-reset |
| Invalid config | exception | info | artifacts, not-found, ttl, expiry |
| Connection failed | exception | error | connection, dns, http-crawler, network |
| invalid header name | validation | error | validation, http-headers, hooks, security |
| pypdf is required for PDF processing. Install with 'pip… | exception | error | pdf, dependency, import-error, installation |
| [NSTProxy] token and channel_id are required | validation | error | proxy, credentials, validation, nstproxy |
| HTTP error for URL | exception | error | extraction, schema-generation, http-status, batch |
| URL blocked | validation | error | crawl4ai, ssrf, webhook, security, network |
| Browser executable not found for type | exception | error | browser, playwright, installation, environment, setup |
| type ' ' may not be constructed from an untrusted request | validation | warning | ssrf-protection, egress, url-validation, relative-url |
| Browser is not available. It may have been closed, crashed… | exception | error | browser-manager, lifecycle, webcrawler |
| URL blocked (SSRF protection) | http | error | crawl4ai, ssrf, http-400, security, network |
| unknown hook action ; allowed | validation | error | validation, hooks, config, api-contract |
| URL must have a valid hostname | validation | error | crawl4ai, validation, webhook, url-parsing |
| The 'fit_html' attribute is deprecated and has been… | exception | warning | crawl4ai, deprecation, content-filter, fit-html |
| C4A script compiler not available. Please ensure… | validation | error | installation, c4a-script, dependencies, import-error |
| Invalid URL format. Must start with http://, https://, or… | http | error | url-validation, http-400, markdown |
| too many headers (max ) | validation | error | validation, pydantic, headers, hooks |
| tool not found | http | error | mcp, http-404, tool-routing, api |
| control characters in value for header | validation | error | validation, security, header-injection, hooks |
| [NSTProxy] Invalid API response — expected a non-empty list | exception | error | proxy, api, schema, nstproxy, retry |
| field ' ' is not permitted on from an untrusted request | validation | warning | ssrf-protection, egress, scheme-validation, url-blocked |
| Invalid URL, make sure the URL is a non-empty string | validation | error | validation, webcrawler, input-validation |
| `domain_or_domains` must be a string or a list of strings. | validation | error | validation, url-seeder, input-validation |
| At least one URL required | http | error | crawl, http-400, validation |
| score_threshold must be between 0.0 and 1.0 | validation | error | validation, configuration, link-preview, threshold |
| Artifact too large | http | error | artifacts, http-413, storage, size-limit |
| Either 'html' or 'url' must be provided | validation | error | extraction, schema-generation, validation |
| timeout must be positive | validation | error | validation, configuration, link-preview, timeout |
| resource not found | http | error | mcp, http-404, resources, api |
| Unsupported filter type | exception | warning | pdf, image-extraction, decoding, data-corruption |
| chunking_strategy must be an instance of ChunkingStrategy | validation | error | validation, chunking, configuration, type-error |
| concurrency must be positive | validation | error | validation, configuration, link-preview, concurrency |
| link_preview_config must be LinkPreviewConfig object or dict | validation | error | validation, link-preview, configuration, type-error |
| Artifact storage quota exceeded | http | error | artifacts, http-507, quota, storage |
| Context files not found | http | warning | crawl4ai, docker, http-404, packaging, llm-context |
| At least one of include_internal or include_external must… | validation | error | validation, configuration, link-preview |
| invalid value for webhook header | validation | warning | crawl4ai, webhook, validation, header-injection, limits |
| max_links must be positive | validation | error | validation, configuration, link-preview |
| too many hooks (max 10) | validation | error | validation, hooks, config, limits |
| Artifact not found | http | error | artifacts, http-404, storage |
| [NSTProxy] Invalid protocol | validation | error | proxy, validation, nstproxy, configuration |
| {e} | exception | error | crawl4ai, css-selector, validation, scraping |
| Invalid limit: . Must be between 1 and 1000 | http | error | http-400, validation, pagination, monitoring |
| invalid webhook header name | validation | warning | crawl4ai, webhook, validation, http-headers |
| (json.dumps) | http | error | |
| too many webhook headers | validation | warning | crawl4ai, webhook, validation, limits |
| Deep crawling with stream currently supports exactly one… | http | error | |
| Hook must return an instance of webdriver.Chrome or None. | validation | error | |
| Script not found in the folder | exception | error | |
| is deprecated | validation | error | |
| must implement 'run(self, url: str, **kwargs) | validation | error | |
| Failed to extract content from the website | exception | error | |
| Regex for ' ' won’t compile after fix | validation | error | |
| Memory at %, refusing new browser | exception | error | |
| artifact exceeds bytes | exception | error | |
| Database path is not set or is empty. | exception | error | |
| LLM did not return valid JSON | exception | error | |
| .run must be async | validation | error | |
| artifact storage quota exceeded | exception | error | |
| Unsupported extraction strategy | exception | error | |
| Failed to crawl | exception | error | |
| Invalid regex for | validation | error | |
| Unsupported chunking strategy | exception | error | |
| Admin scope required | http | error |