workingproxysites.

API directory

Web Data API Directory: Scraping, SERP & Browser APIs

Compare documented scraping APIs, search-result APIs and hosted browsers by interface, output, capabilities and billing basis.

Select the service interface and output needed for a web-data workload.

  • Need returned HTML or extracted data? Compare HTTP APIs.
  • Need an interactive browser session? Compare CDP services and session terms.
  • Need search-engine results? Check the dedicated SERP records and supported parsers.
20 documented listings

This collection covers our reviewed proxy and web-data catalog. It is not an exhaustive market survey or a performance ranking.

Category reviewed . Review due .

Compare documented requirements

20 of 20 listings
Clear filters

Swipe or scroll horizontally to compare requirements. Scroll back to the first column to see listing names.

API interfaces, outputs and billing · alphabetical
API / serviceInterface & outputDocumented behaviorBilling & restrictionsEvidence
Bright Data Browser APIBright Data · official publisher

Hosted browsers for multi-step automation with built-in proxy management and unblocking.

Details & setup →Provider discussionSourced JSON record
CDP / WebSocket

browser session

Hosted browser
Framework support
Official documentation lists Playwright, Puppeteer and Selenium.[1]
Managed infrastructure
The service handles proxy rotation, browser fingerprints, CAPTCHA handling and session recovery.[1]
PAYG rate
The public PAYG plan lists USD 8 per GB with no monthly commitment and built-in proxies and unlocking.[3]
Concurrency
The public pricing page advertises unlimited concurrent sessions.[3]
Setup
Create a Browser API zone and retrieve its connection details from the Overview tab.[2]

Transferred traffic in GB on paid plans; built-in proxy management and unblocking included.

Requirements & limitations
  • Bring-your-own-proxy support, session lifetime and cross-session persistence were not verified.
  • The bandwidth price cannot be directly compared to request-priced HTTP APIs without a workload estimate.
Documentation checkedReview due
3 source references
Bright Data SERP APIBright Data · official publisher

Collect localized search results as parsed JSON or raw HTML, with synchronous and asynchronous delivery.

Details & setup →Provider discussionSourced JSON record
HTTP API

JSON, HTML

SERP API
Result format
Add brd_json=1 to the search URL for parsed JSON; the REST example otherwise returns HTML.[1]
Asynchronous delivery
Enable asynchronous requests on the SERP zone, submit with async=1, then retrieve using the response ID.[1]
Billing
Failed requests are not charged. In asynchronous mode, collection calls are free.[2][1]
PAYG rate
The public PAYG plan lists USD 1.50 per 1,000 requests, without a monthly commitment.[2]

Successful requests; asynchronous result collection is not separately billed.

Requirements & limitations
  • Search-engine and vertical-specific parameters differ; validate the exact parser needed for the workload.
  • Compatibility and delivery behavior are documented; no authenticated request was executed.
Documentation checkedReview due
2 source references
Browserbase Browser SessionsBrowserbase · official publisher

Hosted browser sessions with custom proxies, plan-based concurrency and reconnection through keep-alive.

Details & setup →Sourced JSON record
CDP / WebSocket

browser session

Hosted browser
Custom proxies
Supports custom HTTP/HTTPS proxies, built-in residential proxies and domain-based routing rules.[1]
Location behavior
Built-in proxy geolocation is best effort; the closest available location may be used.[1]
Billing increments
Browser time is billed by the minute and proxy traffic by the MB, with a one-minute and one-MB minimum per session.[2]
Concurrency
Current published limits are 3 sessions on Free, 25 on Developer and 100 on Startup.[2]
Reconnection
A timeout extends duration but a normal session ends on disconnect; keepAlive enables reconnection.[3]

Browser time plus built-in proxy traffic, with plan allowances and overages.

Requirements & limitations
  • Proxy access starts on a paid plan; plan allocations and overage rates differ.
  • Session creation rate is a separate limit from concurrent running sessions.
  • Country and city selection are not strict availability guarantees.
Documentation checkedReview due
4 source references
Browserless Browser as a ServiceBrowserless · official publisher

Connect Playwright or Puppeteer to hosted browsers with optional built-in or third-party proxies.

Details & setup →Sourced JSON record
CDP / WebSocket

browser session

Hosted browser
Time metering
Each started 30-second browser interval costs one unit, including idle time.[1]
Built-in proxy costs
Residential traffic costs 6 units per MB and datacenter traffic 2 units per MB, in addition to browser time.[1]
Custom proxy costs
Third-party proxies configured through --proxy-server do not consume Browserless proxy units; their own provider may bill separately.[1]
Failures
Pre-start validation or queue rejections consume no browser units; failures after startup consume elapsed time and transferred proxy traffic.[1]
Session timeout
The documented default timeout is 30 seconds; a custom timeout can be set on the connection.[1]

Units for browser time, built-in proxy bandwidth and successful CAPTCHA solves.

Requirements & limitations
  • Session duration and concurrency limits depend on the plan.
  • A failed automation run can still consume units.
  • Third-party proxy charges are separate from Browserless browser-time charges.
Documentation checkedReview due
2 source references
Crawlbase Crawling APICrawlbase · official publisher

Fetch page content with normal or JavaScript tokens, optional geo-routing and browser interaction controls.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON

Scraping API
Token types
Normal tokens fetch static responses; JavaScript tokens enable browser rendering. javascript=true can switch a normal-token request to the JavaScript path.[1]
Browser controls
JavaScript requests support page_wait, ajax_wait, scroll and css_click_selector.[1]
Extended scrolling
The first 8 seconds of page-load plus scrolling count as one request; each additional started 5-second interval adds one billed request.[1]
HTTP-method limitation
POST is supported with the normal token only; JavaScript-rendered requests use GET.[1]

Successful requests; JavaScript request configuration and extended scrolling affect billed usage.

Requirements & limitations
  • Exact dollar rates, account concurrency and complete failed-response billing rules remain unverified.
  • Do not assume a long scroll costs the same as one short fetch.
  • SDK compatibility is documented rather than executed.
Documentation checkedReview due
1 source references
Decodo Web Scraping APIDecodo · official publisher

An HTTP API for rendered web content, with request-level choices for proxy pool and JavaScript execution.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON

Scraping API
HTTP integration
The official Dify guide sends POST requests to https://scraper-api.decodo.com/v2/scrape using Basic authentication and a JSON body.[1]
JavaScript rendering
The documented payload uses headless=html to receive HTML after rendering.[1]
Proxy selection
The request can select proxy_pool=premium.[1]
Cost model
The current pricing page separates standard and premium proxies, with and without JavaScript rendering.[2]

Request credits vary by request complexity, including proxy pool and JavaScript rendering.

Requirements & limitations
  • This pilot verifies the v2 HTTP workflow, not the older Advanced-plan documentation.
  • Exact output availability, async behavior, credit costs and failed-request billing should be checked against the selected current plan.
Documentation checkedReview due
2 source references
Nimble Extract APINimble · official publisher

Fetch HTML, Markdown and screenshots with configurable rendering, CSS parsing, browser actions and asynchronous batches.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, markdown, PNG

Scraping API
Formats
The formats array selects HTML, Markdown, Base64 PNG screenshots and response headers.[1]
Rendering
render defaults to false; true enables JavaScript and auto lets Nimble select the configuration.[1]
Parsing and actions
CSS parsing schemas, browser actions and network-capture rules are supported.[1]
Async processing
Async and batch submissions return before processing finishes; poll task or batch progress for completion.[2]
Billing
Extraction Tools are listed at USD 1 per 1,000 URLs, with charges only for successful requests.[3]

Successful URL extractions; public API pricing lists USD 1 per 1,000 URLs for Extraction Tools.

Requirements & limitations
  • The Extraction Tools price is distinct from paid Extraction Templates, Search and Web Search Agent products.
  • Exact concurrency limits and all driver-specific constraints were not verified in this pilot.
Documentation checkedReview due
4 source references
Oxylabs Web Scraper APIOxylabs · official publisher

Fetch public pages or use dedicated target parsers, with rendered HTML and asynchronous batch delivery.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, PNG

Scraping API
Authentication
Use a Web Scraper API user's username and password with HTTP Basic authentication, rather than dashboard credentials.[1]
Rendering and parsing
render=html enables JavaScript rendering. Dedicated supported targets accept parse=true for structured JSON.[1]
Delivery choices
Realtime keeps the connection open; Push-Pull retrieves jobs asynchronously; Proxy Endpoint offers a proxy interface.[2]
Batch size
Push-Pull accepts up to 5,000 query or URL values per batch; Realtime and Proxy Endpoint do not support batch.[2]
Screenshots
render=png returns a Base64-encoded screenshot.[2]

Product-specific plan and target rates; exact costs and failed-request rules were not verified in this pilot.

Requirements & limitations
  • Dedicated parsers apply to supported targets, not every arbitrary URL.
  • Proxy Endpoint accepts fewer additional parameters than the HTTP integrations.
  • Exact plan pricing and failed-request billing remain unverified.
Documentation checkedReview due
2 source references
Scrapeless Agent BrowserScrapeless · official publisher

Connect browser automation frameworks to managed sessions with optional recording, profiles and custom proxies.

Details & setup →Sourced JSON record
CDP / WebSocket

browser session

Hosted browser
Session lifetime
sessionTTL defaults to 180 seconds. The guide recommends a maximum of 900 seconds while noting longer values can be set.[1]
Custom proxies
proxyURL accepts a custom proxy and overrides other proxy parameters; this feature requires a subscription.[1]
Recording
sessionRecording enables replay of a completed session and is disabled by default.[1]
Frameworks and state
Playwright and Puppeteer are supported; Profiles provide login-state reuse with configurable persistence.[2]
Billing dimensions
The Crawl introduction states its runtime-plus-proxy-bandwidth billing model is shared with Browser.[3]

Browser runtime and proxy bandwidth; exact plan-specific rates were not verified.

Requirements & limitations
  • Custom proxies are subscription-only.
  • Page navigation timeout and browser session TTL are separate settings.
  • Country, state and city targeting depend on available proxy infrastructure.
Documentation checkedReview due
3 source references
ScrapeOps Proxy API AggregatorScrapeOps · official publisher

Route scraping requests across upstream services, with rendering, proxy-pool and extraction options.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, markdown, screenshot

Scraping API
Billing outcomes
The credit policy counts both HTTP 200 and 404 responses as successful, billable requests.[1]
Rendering costs
Standard rendering costs 10 credits; residential plus rendering costs 25; render_js_cheap costs 5 using a limited provider set.[1]
Extraction add-on
llm_extract adds 25 credits on top of base proxy costs.[1]
Cheaper rendering tradeoff
render_js_cheap is intended for simpler sites and does not provide the full upstream provider pool.[2]
Response options
The feature list documents screenshot capture and JSON or Markdown LLM extraction responses.[3]

Monthly API credits determined by target and options; HTTP 200 and 404 responses consume credits.

Requirements & limitations
  • Certain domains and premium configurations use different credit schedules.
  • Unused credits do not transfer into the next subscription period.
  • A target 404 can still consume quota.
Documentation checkedReview due
4 source references
ScraperAPIScraperAPI · official publisher

A scraping HTTP API with optional rendering, premium proxy pools and asynchronous requests.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON

Scraping API
Rendering
render=true returns HTML after JavaScript execution; wait_for_selector requires rendering.[1]
Credit examples
The documented parameter table lists 10 credits for rendering, 25 for premium plus rendering, and 75 for ultra-premium plus rendering.[2]
Billable outcomes
HTTP 200 and 404 responses are charged, as are client-cancelled requests when the client allows less than 70 seconds.[2]
Cost inspection
The account/urlcost endpoint estimates request cost and the sa-credit-cost response header reports actual credits.[2]
Concurrency
Concurrency is plan-dependent; the Async API supports batch requests.[3]
Structured extraction
autoparse=true and dedicated structured-data endpoints parse supported sites and target types; ordinary URL fetching is not a universal field-extraction contract.[1][4]

API credits determined by target domain and request options.

Requirements & limitations
  • The listed rendering credits are configuration examples; protected domains and special targets can change total cost.
  • A target 404 can still be billed.
  • Unused subscription credits do not roll over.
Documentation checkedReview due
4 source references
Scrapfly Scrape APIScrapfly · official publisher

A scraping API with managed rendering, JavaScript scenarios and configurable anti-blocking options.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, screenshot

Scraping API
Rendering and actions
render_js enables browser rendering; js_scenario defines page interactions before results return.[1]
Result handling
Examples return scraped content and browser data through the API result rather than exposing a browser connection.[1]
Credit model
A scrape can consume multiple credits depending on browser, proxy and protection configuration.[2]
Concurrency examples
Published monthly plans list 5 concurrent requests for Discovery, 20 for Pro, 50 for Startup and 100 for Enterprise.[2]
Quota expiry
Unused subscription quota does not roll over to the next subscription period.[2]
Integrated extraction
extraction_model, extraction_prompt and extraction_template select documented model, prompt and template extraction approaches within a scrape.[3]

API credits per scrape; proxy network, rendering and protection settings affect consumption.

Requirements & limitations
  • This API listing does not claim Puppeteer or Selenium browser control; the pricing FAQ explicitly distinguishes that capability.
  • Exact target costs and failed-request billing were not verified in this pilot.
Documentation checkedReview due
3 source references
ScrapingAnt Web Scraping APIScrapingAnt · official publisher

Return page HTML through an API with optional browser rendering, selector waits and proxy location controls.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML

Scraping API
Rendering and response
The /v2/general endpoint returns plain HTML. browser defaults to true; return_page_source skips JavaScript rendering while using the browser.[1]
Timeout
Request timeout defaults to 60 seconds and can be set from 5 to 60 seconds.[1]
Credit examples
Simple datacenter requests cost 1 credit; JavaScript-rendered datacenter requests cost 10; rendered residential requests cost 125.[2]
Usage inspection
The Ant-credits-cost response header reports request credit consumption; unused credits expire at the subscription boundary.[2]

API credits based on browser use and proxy network.

Requirements & limitations
  • The credit reference contains an ambiguous intermediate browser/proxy row; this listing includes only clearly identified configurations.
  • Asynchronous operation and failed-request billing were not fully verified.
  • This is an HTTP rendering API, not a CDP browser connection.
Documentation checkedReview due
2 source references
ScrapingBee Google Search APIScrapingBee · official publisher

Fetch Google search results with device, geography and search-type controls.

Details & setup →Provider discussionSourced JSON record
HTTP API

JSON, HTML

SERP API
Authentication
Send the API key as an Authorization Bearer token.[1]
Request mode
light_request defaults to true. Light mode omits the browser and may return less data; disabling it selects regular mode.[1]
Search controls
Select country, desktop/mobile device and search type. News search is not supported with mobile device selection.[1]
HTML and pagination
add_html can include page HTML. The pages parameter supports up to ten pages; the provider recommends at most three for reliability.[1]

Documentation distinguishes light requests at 10 credits from regular requests at 15 credits. Confirm the selected mode; credits are not USD.

Requirements & limitations
  • The general introduction quotes 15 credits while the light-mode section documents a default costing 10.
  • Review pagination and mode charges before estimating a bulk workload.
Documentation checkedReview due
1 source references
ScrapingBee HTML APIScrapingBee · official publisher

Fetch rendered pages, apply CSS extraction rules or run browser scenarios through an HTTP API.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, markdown, text, screenshot

Scraping API
Rendering default
render_js defaults to true. HTML, page text, Markdown and screenshots are documented output options.[1]
Request credits
Classic requests cost 1 credit without rendering or 5 with it; premium with rendering costs 25 and stealth with rendering costs 75.[1]
Extraction and interaction
CSS extract_rules and js_scenario support structured extraction and page interactions.[1]
Custom proxy
The own_proxy parameter supports a user-supplied proxy provider.[1]
Automatic cost control
GET-only mode=auto chooses a successful rendering/proxy tier; max_cost caps eligible tiers and all-failed auto attempts cost zero.[1]

Monthly API credits; rendering and proxy tier determine base credits, with AI extraction add-ons.

Requirements & limitations
  • AI extraction adds credits beyond base fetch costs.
  • Auto mode does not supply page-specific waits or browser scenarios.
  • The service returns extracted content through HTTP; this record does not claim a remotely controllable CDP browser.
Documentation checkedReview due
1 source references
Steel Browser SessionsSteel · official publisher

Create cloud browser sessions with custom proxies, stored browser state and configurable idle release.

Details & setup →Sourced JSON record
CDP / WebSocket

browser session

Hosted browser
Default session
The configuration guide specifies a desktop, headful session with a five-minute timeout; proxies and CAPTCHA solving are off by default.[1]
Custom proxies
Use useProxy.server or proxyUrl; proxyUrl takes precedence over useProxy.[1]
Idle release
inactivityTimeout can release a session before its total timeout; CDP commands and remote input reset the idle timer.[1]
Stored state
sessionContext supplies cookies/storage, while profiles preserve browser state across sessions.[1]
Browser connection
Attach Playwright or Puppeteer using the session ID and API key on the Steel WebSocket endpoint.[2]

Session time; exact plan rates and proxy add-on prices were not verified.

Requirements & limitations
  • The maximum permitted timeout depends on account limits.
  • The current configuration guide lists us-east as the browser region; proxy geography determines the public exit location.
  • Exact concurrency and billing add-ons remain unverified.
Documentation checkedReview due
3 source references
Thordata Universal Scraping APIThordata · official publisher

An HTTP endpoint for HTML or PNG results with optional JavaScript rendering and resource controls.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, PNG

Scraping API
Request interface
The official example posts form-encoded fields to universalapi.thordata.com/request with a Bearer token.[1]
Rendering
js_render=True enables JavaScript rendering for dynamic content.[1]
Output
The type parameter selects HTML or PNG; HTML is the documented default.[1]
Resource controls
block_resources restricts fetched resources; country selects the scraping proxy location.[1]

Not verified; request multipliers, failed-request billing and plan pricing require further evidence.

Requirements & limitations
  • Pricing, failed-request billing, concurrency and asynchronous delivery were not verified.
  • The documented country parameter defaults to no proxy; confirm the selected request configuration.
Documentation checkedReview due
1 source references
Zyte APIZyte · official publisher

Choose HTTP response bodies, rendered HTML, screenshots or supported structured extraction in one API.

Details & setup →Provider discussionSourced JSON record
HTTP API

HTML, JSON, screenshot

Scraping API
HTTP output
httpResponseBody=true returns the response body Base64-encoded; decode it before parsing.[1]
Browser output
Browser requests return rendered HTML, screenshots or both, and support browser actions and network capture.[2]
Structured extraction
The API supports dedicated extraction types including articles, products, job postings and SERPs.[3]
Billing
Unsuccessful and rate-limited responses are free; browser actions, screenshots and extraction can add costs.[4]
Request model
The /v1/extract endpoint blocks until its single-URL result is ready.[3]

Successful responses, priced by target and HTTP/browser tier, with feature add-ons.

Requirements & limitations
  • This listing covers the HTTP extraction API; Zyte's live CDP browser has separate session and billing rules.
  • Initial browser requests do not support arbitrary HTTP methods, bodies or custom headers beyond Referer.
  • Target tier assignments can change; use the provider's cost estimator for a workload.
Documentation checkedReview due
4 source references
Zyte CDP Headless BrowserZyte · official publisher

Drive a live Zyte browser with Playwright or Puppeteer, with session pricing tied to navigation count and time.

Details & setup →Provider discussionSourced JSON record
CDP / WebSocket

browser session

Hosted browser
Connection
The CDP endpoint accepts Playwright, Puppeteer and other CDP-compatible libraries.[1]
Session lifetime
ttl defaults to 60 seconds and accepts values from 15 to 3,600 seconds.[1]
Billing
Session units are the larger of navigation count and duration rounded up to 15-second intervals, using the first target's browser price tier.[1]
Continuity
A connection has its own IP and cookie jar; separate CDP connections cannot reuse that session automatically.[1]
Access requirements
CDP requires a subscription or PAYG spending limit plus business verification; free trials cannot use it.[1]

Target-tier units: the greater of page.goto calls or started 15-second session intervals; residential overages can add charges.

Requirements & limitations
  • Started sessions can be billed even if the automation fails.
  • Zyte documents that Playwright browser.close() disconnects without sending the Browser.close CDP command; follow its explicit closing example.
  • Bring-your-own-proxy support was not verified.
  • The official documentation conflicts on whether explicit Browser.close stops billing before ttl. Budget the configured ttl conservatively until Zyte confirms the applicable behavior.
Documentation checkedReview due
1 source references
Zyte Search APIZyte · official publisher

Request parsed search results or rendered result-page HTML through a dedicated search endpoint.

Details & setup →Provider discussionSourced JSON record
HTTP API

JSON, HTML

SERP API
Endpoint and authentication
POST to /v1/search using the API key as the Basic-auth username and an empty password.[1]
Output selection
The include option selects organic results, HTML or both. Parsed results carry rank, title and URL; snippet is optional.[3][1]
Result quantity
maxResults accepts multiples of ten from 10 to 100, with a default of 10. Request weight increases with result count.[2]
Targeting
Use engine-native query parameters or generic geolocation and locale options. Unsupported search domains return HTTP 400.[2]

Request weight increases with maxResults: max(1, maxResults / 10). A USD rate was not established from the reviewed search documentation.

Requirements & limitations
  • Parsed AI Overview extraction is not yet available. The quickstart describes returning its rendered HTML, while the parameter table still labels aiOverview coming soon; confirm this mode before depending on it.
  • Search-specific USD pricing and account limits were not established.
Documentation checkedReview due
3 source references

How to use this comparison

Compare the documented request interface, output and account product before checking price. A JSON response can contain raw HTML; it does not by itself establish structured extraction. A browser action in one request does not by itself establish a reusable browser session.

Each row keeps the publisher’s billing basis and restrictions alongside its sources. Unknown prices or limits remain unknown. These are documentation reviews; no live connection, purchase or comparative performance test is claimed.

Overdue records remain visible with a label. This collection leaves search indexing when its category review is overdue or fewer than 3 current, visible records qualify. Individual detail pages retain their own review policy.

The collection JSON link covers the whole category before filters.

Suggest a correction · Affiliate disclosure · Collection JSON records