API directory
Structured Web Extraction APIs: Compare Parsers & Output Contracts
Compare scraping APIs with documented structured-extraction capabilities, outputs, authentication requirements and billing qualifications.
Shortlist extraction APIs and distinguish parsed target data from a JSON response envelope.
- Which target types or extraction rules are supported?
- Is JSON parsed target data, or only a wrapper containing HTML?
- Do custom fields, browser rendering and asynchronous delivery change the request contract?
This collection covers our reviewed proxy and web-data catalog. It is not an exhaustive market survey or a performance ranking.
Category reviewed . Review due .
Compare documented requirements
7 of 7 listingsSwipe or scroll horizontally to compare requirements. Scroll back to the first column to see listing names.
| API / service | Interface & output | Documented behavior | Billing & restrictions | Evidence |
|---|---|---|---|---|
| Nimble Extract APINimble · official publisher Fetch HTML, Markdown and screenshots with configurable rendering, CSS parsing, browser actions and asynchronous batches. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, markdown, PNG Scraping API | For this task: A CSS parsing schema defines extracted fields separately from the formats array. Match selectors to the target page and choose rendering independently; requesting a JSON response alone does not define a field schema.[1][4]
| Successful URL extractions; public API pricing lists USD 1 per 1,000 URLs for Extraction Tools. Requirements & limitations
| Documentation checkedReview due 4 source references
|
| Oxylabs Web Scraper APIOxylabs · official publisher Fetch public pages or use dedicated target parsers, with rendered HTML and asynchronous batch delivery. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, PNG Scraping API | For this task: Dedicated supported targets accept parse=true for parsed fields. A universal page fetch does not establish automatic parsing for every site; check the supported source and target before choosing a schema.[1]
| Product-specific plan and target rates; exact costs and failed-request rules were not verified in this pilot. Requirements & limitations
| Documentation checkedReview due 2 source references
|
| ScrapeOps Proxy API AggregatorScrapeOps · official publisher Route scraping requests across upstream services, with rendering, proxy-pool and extraction options. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, markdown, screenshot Scraping API | For this task: The LLM extraction path supports JSON output and page-type schema hints. llm_extract adds 25 credits to the underlying fetch cost; this is separate from receiving an ordinary JSON response wrapper.[4][1]
| Monthly API credits determined by target and options; HTTP 200 and 404 responses consume credits. Requirements & limitations
| Documentation checkedReview due 4 source references
|
| ScraperAPIScraperAPI · official publisher A scraping HTTP API with optional rendering, premium proxy pools and asynchronous requests. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON Scraping API | For this task: autoparse=true and dedicated structured-data endpoints work with supported sites and target types. An ordinary URL fetch does not promise a general parser for arbitrary websites.[1][4]
| API credits determined by target domain and request options. Requirements & limitations
| Documentation checkedReview due 4 source references
|
| Scrapfly Scrape APIScrapfly · official publisher A scraping API with managed rendering, JavaScript scenarios and configurable anti-blocking options. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, screenshot Scraping API | For this task: Integrated extraction accepts extraction_model, extraction_prompt or extraction_template. These select model, prompt and template approaches with different output contracts; choose the mechanism and required fields for the target.[3]
| API credits per scrape; proxy network, rendering and protection settings affect consumption. Requirements & limitations
| Documentation checkedReview due 3 source references
|
| ScrapingBee HTML APIScrapingBee · official publisher Fetch rendered pages, apply CSS extraction rules or run browser scenarios through an HTTP API. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, markdown, text, screenshot Scraping API | For this task: CSS extract_rules defines the fields to extract. AI extraction rules are a separate option with additional credit requirements; choose the extraction mechanism before comparing costs.[1]
| Monthly API credits; rendering and proxy tier determine base credits, with AI extraction add-ons. Requirements & limitations
| Documentation checkedReview due 1 source references
|
| Zyte APIZyte · official publisher Choose HTTP response bodies, rendered HTML, screenshots or supported structured extraction in one API. Details & setup →Provider discussionSourced JSON record | HTTP API HTML, JSON, screenshot Scraping API | For this task: Choose a typed extraction output such as product or article. Returned HTML and a JSON response envelope are separate from these extracted objects; select the page type and extraction source for the workload.[3]
| Successful responses, priced by target and HTTP/browser tier, with feature add-ons. Requirements & limitations
| Documentation checkedReview due 4 source references
|
How to use this comparison
Compare the documented request interface, output and account product before checking price. A JSON response can contain raw HTML; it does not by itself establish structured extraction. A browser action in one request does not by itself establish a reusable browser session.
Each row keeps the publisher’s billing basis and restrictions alongside its sources. Unknown prices or limits remain unknown. These are documentation reviews; no live connection, purchase or comparative performance test is claimed.
Overdue records remain visible with a label. This collection leaves search indexing when its category review is overdue or fewer than 3 current, visible records qualify. Individual detail pages retain their own review policy.
The collection JSON link covers the whole category before filters.
Suggest a correction · Affiliate disclosure · Collection JSON records