Managed service

A public records API for your devs. RPA behind it. Supported by us.

Your product team wants an endpoint, not a data project. Recordpipe stands up a custom API on exactly the dataset you need: your schema, REST endpoints, keys and documentation your developers can build against in an afternoon — while our managed automation keeps the data behind it collected, normalized, and current.

What we deliver

Your schema, your endpoints

The API is designed in scoping around how your product queries — not around how the source data happens to look.

Managed collection behind it

RPA pipelines refresh the underlying data daily, weekly, or as-needed. Sources change; we fix it; your endpoint doesn't flinch.

Keys, docs, versioning

Standard REST with API keys, written docs, and versioned changes — the things your devs expect from a real vendor.

Webhook events included

Beyond request/response: subscribe your systems to pushes when records matching your criteria appear or change.

What ships with a custom data API

A custom data API from Recordpipe is a finished integration surface, not a database dump with a URL. The build includes:

  • Endpoints designed around your product's queries — search, single-record fetch, history, and a change feed — shaped the way your developers would have designed them.
  • API keys with per-key scoping, separating staging, production, and internal tooling from day one.
  • Written documentation with example requests and responses your team can build against without a kickoff call.
  • A versioned schema with a published change policy — additive changes flow in, breaking changes get a new version and a migration window.
  • Webhook subscriptions for push delivery when records matching your criteria appear or change.

Staging keys are live during the build, so your integration and our pipeline are tested against each other before launch, not after.

Data as a service API: the pipeline is the product

The endpoint is the visible part. What makes this a data as a service API rather than a hosted spreadsheet is the managed collection layer behind it. On the cadence set in scoping, our automation gathers the underlying public records, normalizes them to your schema, deduplicates entities, and computes what changed — and those deltas are what feed the change endpoint and your webhooks.

Every record carries provenance: a capture date and a source category, so your product can show users how current a value is and your auditors can trace where it came from. The stance underneath doesn't move: publicly available sources only — nothing login-walled, nothing private, no breached data, under any contract.

Because collection and serving are one system, freshness is observable rather than promised: a status endpoint reports the last completed refresh per dataset, for your monitoring to alert on.

A managed scraping API for developers — not a scraping tool with an API

Most products sold as a scraping API take a URL and return raw page content. Your team still writes the parsers, still maintains them when pages change, and still owns data quality — the hard part never left your backlog. A managed scraping API for developers inverts that: you never send a URL, and you never see a page. You query records.

Extraction, parsing, validation, and repair all run on our side as a managed web scraping service, on the same production infrastructure that processes over one million public records nightly. Your integration surface is stable JSON with a contract we version. When a source restructures, we fix the pipeline behind the endpoint; your code doesn't know it happened.

Your developers' time goes into your product's features. The part that decays — source maintenance — sits with the vendor built to absorb it.

From scoping call to production keys

The path to a live endpoint is short and priced up front. Describe the dataset and how your product needs to query it; we reply within 2 business days with feasibility and an approach. The $500 scoping — credited in full against the contract — delivers a data sample and a proposed endpoint design within 5 business days, so your engineers review real JSON, not a slide.

From there: we finalize the schema together, stand up the pipeline and staging environment, your team integrates against staging keys, and launch happens when both sides sign off. The build is fixed-price — contracts run from $5,000 to $3 million depending on dataset scope and refresh cadence — and ongoing operation is a retainer quoted alongside it, with no metered-billing surprise.

When a custom API becomes a product

Some datasets stop being an internal dependency and start being a feature you sell. We build for that path from the start. Keys are multi-tenant by design, usage is metered per key, and the documentation is written cleanly enough to hand to your own customers. If your roadmap includes exposing the data inside your product — an in-app lookup, a verification step, an alerts feed — the API you commissioned is already shaped for it.

The commercial side scales the same way: licensing terms for downstream use are set in your contract, not renegotiated per customer you sign. Teams that later ship the endpoint as a product feature do it on the same build and the same pipeline, with capacity resized in the retainer rather than re-architected.

An example endpoint spec

EndpointMethodDescription
/v1/records/searchGETQuery the dataset by the parameters your product actually filters on — name, jurisdiction, status, date range. Paginated, with stable sort.
/v1/records/{id}GETFetch a single record in full, including capture date and source category for every field group.
/v1/records/{id}/historyGETPrior observed values for a record, so your product can show what changed and when.
/v1/changesGETCursor-based change feed: everything added, updated, or no longer present at source since your last cursor.
/v1/webhooksPOSTRegister a push subscription — your criteria, your endpoint, signed payloads.
/v1/webhooks/{id}DELETERemove a subscription. Deliveries stop immediately; the audit log of past pushes remains queryable.
/v1/exportsPOSTStart a bulk export job — full dataset or a filtered slice — in CSV, JSON, or Parquet.
/v1/exports/{id}GETPoll export status and retrieve the download link when the job completes.
/v1/statusGETPipeline health and last completed refresh per dataset, for your monitoring to alert on.
How teams use it

In the field.

In-app property lookup at a proptech company

A proptech product team can ship an ownership-lookup feature backed by a records endpoint on our infrastructure: their app queries by address, renders the public ownership and transaction history, and shows the capture date. Their engineers integrate an API; they never touch a records pipeline.

Merchant onboarding at a payments company

A payments company's onboarding team can call a business-registration endpoint at signup: entity status, registered agents, and filing history returned as structured JSON, feeding their own risk logic. The change feed then flags dissolutions or status changes across the existing merchant base.

Quote-flow verification at an insurtech

An insurtech building contractor coverage can verify license status inside the quote flow — one API call per applicant, answered from data our pipelines keep refreshed. Webhooks notify the book-management team when an insured contractor's public license status changes mid-policy.

An alerts feature at a legal-tech startup

A legal-tech team can build a monitoring feature on the change feed and webhooks: when new public filings matching a client's watch criteria appear, the app raises an alert linking to the full record. The startup sells the feature; we run the collection behind it.

Government data API, bulk data API, data feed API: from custom endpoint to product at scale

When one dataset keeps getting requested, an API serving many customers beats custom builds serving one — we build for that path from day one. Contracts are fixed-price, $5,000 to $3 million by scope and cadence; scoping is $500, credited in full.

Who hosts and maintains the API?
We do — the endpoints, the keys, the monitoring, and the collection pipelines behind them. "Supported by us" means your team never operates any of it: when a source changes, we repair the pipeline and your endpoint keeps answering.
What does it cost to keep running?
The build is fixed-price; ongoing refresh and hosting is a retainer quoted with your contract, sized to cadence and volume. No per-request metering surprises — the retainer is the number.
How do you handle breaking changes?
The schema is versioned. Additive fields appear without ceremony; anything that would break your integration ships as a new version with a documented migration window, and the old version keeps answering until you've moved.
What are the rate limits?
Throughput is sized to your contract during scoping, based on how your product will call the API — there's no shared public tier to compete with. If usage grows, capacity is resized in the retainer, not throttled.
Can we get bulk data alongside the API?
Yes. Export endpoints produce full or filtered extracts in CSV, JSON, or Parquet, and scheduled file drops to your bucket or warehouse can run alongside request/response access on the same pipeline.
Can we expose the data inside our own product?
Usually, yes — downstream licensing is set in your contract up front. The API is built multi-tenant with per-key metering, so an internal endpoint can become a customer-facing feature without a rebuild.
Can we use it for tenant or employment screening?
What the API serves is raw public data, not consumer reports — Recordpipe is not a CRA. Eligibility use cases get a compliance review at intake so the boundaries are settled in the design, not discovered in production.