Web scraping use cases and recipes
Web data API use cases with runnable recipes, measured credit costs and honest limits: job postings, protected sites, AI knowledge bases, leads, prices.
Each use case starts from the data you want and ends with a runnable request, its credit cost and what does not work yet.
Which Spicrawl use case should I start with?
Scrape job postings
Job boards, infinite-scroll listings and public job APIs: 15 of 16 sources worked, with request and cost.
Scrape bot-protected sites
Cloudflare and PerimeterX cleared on your own proxy; blocked Akamai, Kasada and Imperva cost 0.
Turn websites into LLM-ready data
Clean markdown for RAG and AI agents: batch ingestion, chunking, MCP and scheduled refreshes.
Lead generation from the web
Company directories and listing sites into structured leads, exported to CSV or a CRM.
E-commerce price monitoring
Product price, availability and rating: re-scrape fresh, batch catalogs, compare countries.
Which Spicrawl guides do the use cases build on?
JavaScript rendering
Run a page in a real browser with js_render for 3 credits.
Browser actions
Scroll, click and fill before the page is captured.
Structured data
Get fields as JSON with autoparse or selectors.
Batch jobs
Fetch up to 10,000 URLs in one server-side job.
Proxies and geo
Route through your own proxy and match the exit country.
Sessions and logins
Log in once and reuse the cookies.
Start scraping in minutes
1,000 free credits every month. One API key, one request.
Disclaimer
For educational purposes
The examples on this page are for educational purposes only; the URLs are placeholders, and Spicrawl is not affiliated with any site you scrape. Check each site's terms and robots.txt, respect rate limits, and follow the laws that apply to you, including data-protection laws such as the GDPR and CCPA when pages contain personal data. You are responsible for how you use Spicrawl and the data you collect. This is not legal advice.
Caching
How the /v1/scrape result cache works: on by default with a 48-hour window, a cache hit is billed at the price of the fetch that stored it, and cache=false or cache_ttl=0 forces a fresh fetch.
Job postings
Job scraping API cookbook: scrape job boards, infinite-scroll listings and public job APIs with Spicrawl. Tested requests, credit costs and limits.