spicrawlspicrawlDocs
Batch

Get one item's payload

Returns a single finished item's payload verbatim (not wrapped), capped at 512 KiB.

Requires scope: batch

GET
/v1/batch/{batchID}/tasks/{seq}/content

Returns a single finished item's payload verbatim (not wrapped), capped at 512 KiB. Use it to fetch one URL's output without paging the whole result set. Status meanings: 200 payload; 409 either "not finished yet, poll the job and ask again" or "finished as failed/cancelled/skipped, will never have content" (read detail; for failures, POST .../retry); 404 seq is beyond the job's item count; 503 the deployment has no per-item payload to serve. Note: the worker currently writes one stitched file per job and no per-item reference, so this endpoint answers 503 for succeeded items in practice — read result.content from the results endpoint instead. Retention is not checked here.

Authorization

bearerAuth
AuthorizationBearer <token>

Authorization: Bearer <key>. Read the key from the SPICRAWL_API_KEY environment variable; never hard-code or log it. spicrawl_test_… keys can never spend live credits. Scopes: scrape, batch, sessions (granted by default), browser and read (granted deliberately). A missing scope is 403 ERR::AUTH::INSUFFICIENT_SCOPE naming the scope.

In: header

Path Parameters

batchID*string

The job id (26-character ULID as returned). The 32/36-character UUID form of the same id is also accepted. A malformed id and another tenant's id both answer 404.

Length26 <= length <= 36
seq*integer

The item's 0-based position in submission order (seq on a result line). Non-integer or negative is a 400.

Range0 <= value

Response Body

application/json

application/problem+json

application/problem+json

application/problem+json

application/problem+json

application/problem+json

application/problem+json

application/problem+json

curl -X GET "https://example.com/v1/batch/01J9Z6T3W9E21T5TZARVJRVN5C/tasks/0/content"
null