Data API / DataForSEO / On Page
Uncrawlable Resources
The resources a crawl could not fetch, and why. Returns total_items_count, items_count and items. Reads a finished crawl, so it needs the id from post_dataforseo_on_page_submit and returns crawl_progress plus a crawl_status of max_crawl_pages, pages_in_queue and pages_crawled - check those before trusting a small result, because a crawl still running simply has less to report. These are the broken references a site owner would want first; the ones that did load are in post_dataforseo_on_page_resources. 💰 Free upstream: querying a finished crawl costs nothing, only the crawl itself does.
Параметры
Параметров пути, запроса и заголовков нет.
Тело запроса
idstringОбязательныйID of the task required field you can get this ID in the response of the Task POST endpoint example: "07131248-1535-0216-1000-17384017ad04"
limitintegerНеобязательныйthe maximum number of returned uncrawlable resources optional field default value: 100 maximum value: 1000
offsetintegerНеобязательныйoffset in the results array of returned uncrawlable resources optional field default value: 0 if you specify the 10 value, the first ten invalid resources in the results array will be omitted and the data will be provided for the successive invalid resources
order_bystring[]Необязательныйresults sorting rules optional field you can use the same values as in the filters array to sort the results possible sorting types: asc - results will be sorted in the ascending order desc - results will be sorted in the descending order you should use a comma to set up a sorting type example: ["meta.content_type,desc"] note that you can set no more than three sorting rules in a single request you should use a comma to separate several sorting rules example: ["meta.content_type,asc","fetch_time,desc"]
filtersarrayНеобязательныйarray of results filtering parameters optional field you can add several filters at once (8 filters maximum) you should set a logical operator and, or between the conditions the following operators are supported: regex, not_regex, , , >, >=, =, , in, not_in, like, not_like you can use the % operator with like and not_like to match any string of zero or more characters example: [["meta.content_type","=","image/jpeg"], "and", ["url","not_like","%/help-center/%"]]The full list of possible filters is available by this link.
Пример запроса
curl -X POST "https://openapi.felo.ai/v1/beta/dataforseo/on_page/uncrawlable_resources" \
-H "Authorization: Bearer $FELO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"id": "<string>",
"limit": 0,
"offset": 0,
"order_by": [
"<string>"
],
"filters": "<string>"
}'Ответ
Поля ответа
versionstringНеобязательныйthe current version of the API
status_codeintegerНеобязательныйgeneral status code you can find the full list of the response codes here Note: we strongly recommend designing a necessary system for handling related exceptional or error conditions
status_messagestringНеобязательныйgeneral informational message you can find the full list of general informational messages here
timestringНеобязательныйexecution time, seconds
costnumberНеобязательныйtotal tasks cost, USD
tasks_countintegerНеобязательныйthe number of tasks in the tasks array
tasks_errorintegerНеобязательныйthe number of tasks in the tasks array returned with an error
tasksstring[]Необязательныйarray of tasks
idstringНеобязательныйtask identifier unique task identifier in our system in the UUID format
status_codeintegerНеобязательныйstatus code of the task generated by DataForSEO; can be within the following range: 10000-60000 you can find the full list of the response codes here
status_messagestringНеобязательныйinformational message of the task you can find the full list of general informational messages here
timestringНеобязательныйexecution time, seconds
costnumberНеобязательныйcost of the task, USD
result_countintegerНеобязательныйnumber of elements in the result array
pathstring[]НеобязательныйURL path
dataobjectНеобязательныйcontains the same parameters that you specified in the POST request
resultstring[]Необязательныйarray of results
crawl_progressstringНеобязательныйstatus of the crawling session possible values: in_progress, finished
crawl_statusobjectНеобязательныйdetails of the crawling session
max_crawl_pagesintegerНеобязательныйmaximum number of pages to crawl indicates the max_crawl_pages limit you specified when setting a task
pages_in_queueintegerНеобязательныйnumber of pages that are currently in the crawling queue
pages_crawledintegerНеобязательныйnumber of crawled pages
total_items_countintegerНеобязательныйtotal number of uncrawlable resources found total number of uncrawlable resources found during the crawl of the target domain
items_countintegerНеобязательныйnumber of uncrawlable resources in the items array
itemsstring[]Необязательныйarray of uncrawlable resources
urlstringНеобязательныйURL of the uncrawlable resource
reasonstringНеобязательныйreason the resource is uncrawlable can take the following values: content_type_inconsistency
status_codeintegerНеобязательныйHTTP response code returned by the uncrawlable resource possible values: 200
fetch_timestringНеобязательныйdate and time when the resource was fetched in the UTC format: “yyyy-mm-dd hh-mm-ss +00:00” example: 2026-03-09 18:20:32 +00:00
metaobjectНеобязательныйmetadata of the uncrawlable resource
content_typestringНеобязательныйactual content type of the resource
expected_content_typesstring[]Необязательныйexpected content types for the resource list of content types that were expected by the crawler based on how the resource is referenced on the page