Priced by bandwidth, not requests.
Retrieval is billed by bandwidth. Search uses credits; structured extraction uses both. Charges depend on the endpoint and options.
Turn public pages into Markdown or structured data for your agents and pipelines. ByteKit handles stealth, fetching, browser rendering, and extraction.
npm install @hunt-labs/bytekit-sdk pip install bytekit-sdkRetrieval is billed by bandwidth. Search uses credits; structured extraction uses both. Charges depend on the endpoint and options.
Scrape cache hits halve the bandwidth charge. Choose your cache window, or fetch fresh content.
Carry unused monthly bandwidth forward with no cap while you keep your plan or upgrade.
Get the web data your application needs without building the retrieval infrastructure.
Browser rendering, content conversion, caching, and scheduled checks—without running the infrastructure yourself.
ByteKit can fall back to a browser when an HTTP fetch isn’t enough. Set wait conditions for content that loads later.
Request Markdown, raw HTML, or structured fields. Capture screenshots or discover URLs through sitemaps.
Find URLs with search, then retrieve the pages you choose. Already have a URL? Start with retrieval.
Configure a monitor and webhook. ByteKit re-fetches the page and notifies your application when content changes.
Scrape cache hits cost 50% less in bandwidth charges. Choose your cache window, or turn caching off for a fresh fetch.
Retrieve content with available page metadata, such as title, language, and description, for your application to use.
Give agents context, build datasets, or track changes on public pages.
Let agents retrieve public pages as Markdown, with search available when needed.
Retrieve Markdown and page metadata for your chunking, embedding, and indexing pipeline.
Capture pages on a schedule for your training and evaluation datasets.
Monitor public pricing and product pages. Receive a webhook when page content changes.
Extract prices, specifications, and availability from public product pages into fields you define.
Extract industry, product, and location details from public company pages for account research.
Keep your application logic. Hand off anti-bot handling, browser infrastructure, extraction, and monitoring.
| Responsibility | Building it yourself | With ByteKit |
|---|---|---|
| Anti-bot handling | Maintain compatibility as anti-bot protections change. | ByteKit maintains anti-bot handling. |
| Browser infrastructure | Operate browsers and manage rendering, failures, and retries. | A simple API request. We choose the retrieval strategy. |
| Extraction & formatting | Maintain parsing and conversion logic. | Request Markdown or define fields with a JSON schema. |
| Keeping data current | Schedule checks, detect changes, and deliver notifications. | Configure a monitor and webhook. |
Use ByteKit for public pages you’re authorized to access.
Login workflows and paywalled content aren’t supported.
We block .gov domains and selected financial, healthcare, ticketing, and social media sites.
No ticket-purchase bots, queue jumping, reservation bots, or inventory hoarding.
No doxxing, stalking, or unauthorized tracking of individuals.
Phishing, impersonation, scams, and malware-related use are prohibited.
Technical access isn’t permission. Respect restrictions and requests to stop.
New accounts receive a one-time allowance of 100 MB and 100 credits.