sdkProvider-backed

Scrape website

pipeline.scrapeWebsite

Pipeline recommended website scraping via workspace Spider.cloud BYOK. Collect markdown using scrape or bounded crawl. No AI extraction, platform key, or automatic paid analysis. Use public URLs. Returns source pages for table research.

Parameters

workspaceIdrequired
string
No description.
credentialId
string
No description.
urlrequired
string
No description.
mode
string
No description.
scrape | crawl
limit
integer
No description.
depth
integer
No description.

Example

TypeScripttypescript
await pipeline.scrapeWebsite({
"workspaceId": "00000000-0000-4000-8000-000000000000",
"url": "https://example.com"
})

Response

Success response schemajson schema
{
"type": "object",
"properties": {
"success": {
"type": "boolean"
},
"provider": {
"type": "string",
"enum": [
"spider"
]
},
"pages": {
"type": "array",
"items": {
"type": "object",
"properties": {
"url": {
"type": "string"
},
"content": {
"type": "string"
}
}
}
}
},
"required": [
"success",
"provider",
"pages"
]
}

Exceptions

PipelineError
Thrown for validation, authentication, JSON-RPC, HTTP, and network failures. Inspect code and data when present.
Pipeline TypeScript SDK documentation | Pipeline