資料 API / DataForSEO / On Page
OnPage API Raw HTML
The stored HTML of one crawled url. Reads a finished crawl, so it needs the id from post_dataforseo_on_page_submit and returns crawl_progress plus a crawl_status of max_crawl_pages, pages_in_queue and pages_crawled - check those before trusting a small result, because a crawl still running simply has less to report. ⚠️ Only available if the crawl was submitted with store_raw_html set - otherwise there is nothing to return, and the crawl cannot be amended after the fact. For parsed content rather than source use post_dataforseo_on_page_content_parsing. 💰 Free upstream: querying a finished crawl costs nothing, only the crawl itself does.
參數
沒有路徑、查詢或請求標頭參數。
請求主體
idstring必填ID of the task required field you can get this ID in the response of the Task POST endpoint example: “07131248-1535-0216-1000-17384017ad04”
urlstring選填page url required field the absolute URL of a page to request HTML Note: this field is optional if the task was set using the Instant Pages endpoint
請求範例
curl -X POST "https://openapi.felo.ai/v1/beta/dataforseo/on_page/raw_html" \
-H "Authorization: Bearer $FELO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"id": "<string>",
"url": "<string>"
}'回應
回應欄位
versionstring選填the current version of the API
status_codeinteger選填general status code you can find the full list of the response codes here Note: we strongly recommend designing a necessary system for handling related exceptional or error conditions
status_messagestring選填general informational message you can find the full list of general informational messages here
timestring選填execution time, seconds
costnumber選填total tasks cost, USD
tasks_countinteger選填the number of tasks in the tasks array
tasks_errorinteger選填the number of tasks in the tasks array returned with an error
tasksstring[]選填array of tasks
idstring選填task identifier unique task identifier in our system in the UUID format
status_codeinteger選填status code of the task generated by DataForSEO; can be within the following range: 10000-60000 you can find the full list of the response codes here
status_messagestring選填informational message of the task you can find the full list of general informational messages here
timestring選填execution time, seconds
costnumber選填cost of the task, USD
result_countinteger選填number of elements in the result array
pathstring[]選填URL path
dataobject選填contains the same parameters that you specified in the POST request
resultstring[]選填array of results
crawl_progressstring選填status of the crawling session possible values: in_progress, finished
crawl_statusobject選填details of the crawling session
max_crawl_pagesinteger選填maximum number of pages to crawl indicates the max_crawl_pages limit you specified when setting a task
pages_in_queueinteger選填number of pages that are currently in the crawling queue
pages_crawledinteger選填number of crawled pages
items_countinteger選填number of items in the results array
itemsobject選填items object
htmlstring選填HTML page