Some links on this page are affiliate links: if you buy through them we may earn a commission, at no extra cost to you.
A browser API supplies a remote browser; custom rules tell it what to do on a particular website. Together, they can load JavaScript-driven content, interact with page controls, and return a result for extraction—but the rules must be tailored to the site and checked as it changes.
What custom rules add to a browser API
A browser API is the execution environment, not a ready-made scraper for every website. Custom rules describe the page-specific navigation and interactions needed to expose the data you want. For example, they might open a page, enter a search phrase, click a control, wait for results, and return the resulting HTML or structured data.
Oxylabs describes this pattern as submitting instructions for a target page, having a browser execute them, and transferring the result—raw HTML or structured JSON—to storage. Its Custom Browser Instructions also explain why this can help with JavaScript-heavy components: the initial response may not contain content that appears only after the page renders or an interaction triggers more loading.
How the workflow works
- Inspect the page. Identify the data you need and the elements or controls that display or reveal it. A search box, dropdown, “load more” button, or result container may be part of the path.
- Write the interaction sequence. Specify actions such as filling a field, clicking a control, scrolling, or running JavaScript. Add a wait condition that corresponds to the result you need, if the service supports it.
- Let the browser reach the required state. The remote browser runs those instructions against the page. JavaScript may then make additional requests and insert their results into the document.
- Retrieve and parse the result. The service may return HTML or structured JSON. Check that the expected fields are present and have plausible values rather than assuming a successful browser run guarantees correct extraction.
- Test against the actual site. Validate selectors, waits, and navigation on the target page before relying on the workflow. Web Scraper’s documentation cautions that no universal tool can guarantee compatibility with every website.
When browser automation is worth using
Use it when page state depends on JavaScript or interaction
A browser-based workflow is useful when data appears only after rendering, a click, a form fill, a dropdown selection, scrolling, or waiting for an element. It can also suit an existing Puppeteer, Playwright, or Selenium workflow that needs a managed remote browser.
#1 Best Overall
Prefer a lighter method for simple HTTP retrieval
If a page’s required content is available from a straightforward HTTP request and needs no browser interaction, full browser automation may add unnecessary overhead. Bright Data’s Browser API reference distinguishes simple HTTP scraping from browser tasks such as clicking, filling forms, running JavaScript, handling single-page applications, or intercepting page XHR/fetch requests. That is vendor guidance, not a universal performance comparison.
Choose an approach by workflow, not by a promised winner
| Approach | What it does | What to compare |
|---|---|---|
| Custom-instruction scraping API | Runs website-specific browser actions and returns HTML or structured JSON. | Required actions, output format, wait behavior, maintenance, and service price. |
| Framework-connected cloud browser | Connects Puppeteer, Playwright, or Selenium to a managed browser. | Framework support, session setup, control, debugging, and operational complexity. |
| Sitemap-based extension or cloud scraper | Defines navigation and data selectors in a sitemap; hosted options may add scheduling and delivery. | Local versus hosted execution, selector validation, scheduling, retries, and export. |
| Trained-agent scraper | Configures an agent to capture named fields and invokes it through an API, webhook, or polling workflow. | Setup effort, field structure, resilience to page changes, and integration. |
These are different operating models, not evidence of a benchmark winner. Compare candidates using the same target pages, required output, and current plan details. For an interaction-based workflow, Scrape.do’s Browser Interactions documentation describes actions, waits, and per-action success or error information; Browse AI describes trained-agent extraction and API, webhook, or polling workflows in Turn any website into an API.
Rank #2
Catch failures before they become bad data
- A selector no longer matches: the page layout or control may have changed. Check action-level errors where available and validate the rules against the live target.
- Extraction starts too soon: a fixed delay can expire before the relevant content arrives. Prefer waiting for the target element or request when supported, then verify the returned fields.
- The interaction behaves differently: controls and navigation can vary by browser mode. Scrape.do notes that in its Android-based mobile browser infrastructure, taps require its Tap action because Click does not work there.
- The page changes over time: selectors and steps need maintenance. Treat target-site validation and monitoring as part of operating the scraper, not as a one-time setup.
Or skip the browser setup
If what you need is a screenshot or PDF rather than extracted page data, ScreenshotNeo is a website screenshot API and MCP server—not a general-purpose structured-data scraper. Its API can return a screenshot in PNG, JPEG, or WebP, or a PDF. One request:
Quick Recap
Best Value
curl -G "https://api.screenshotneo.com/v1/shot" -d access_key=YOUR_API_KEY --data-urlencode url=https://stripe.com -o shot.webp
See the ScreenshotNeo API documentation for request options. Before capture, it can accept cookie or consent banners and remove more than 60 known consent platforms, newsletter popups, and chat widgets; each of those steps can be turned off. Bot checks or CAPTCHAs, blank pages, timeouts, failed loads, and cache hits cost nothing, and response headers identify the page verdict and billing status. Its MCP server offers take_screenshot, get_page_info, and capture_pdf for AI agents. The free plan includes 1,000 screenshots per month without a card; paid plans start at $5 for 3,000.
Sign up for ScreenshotNeo’s free plan.
Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.

