Solutions

A data supplier for the data industry.

Data companies know exactly what to sell; the grind is collecting and normalizing it, county by county, source by source, forever. Recordpipe runs that layer as a service: we build the pipeline, you own the output — bulk corpora, ongoing refresh operations, or an entire dataset category you've wanted in your catalog but never had the collection capacity to build.

What we deliver

Multi-jurisdiction corpus builds

Hundreds of sources normalized to one schema in a single fixed-price program.

Refresh operations

We run the ongoing collection on daily/weekly cadence; you get clean deltas into your existing catalog format.

Gap-filling

The counties and record types your current suppliers can't or won't cover.

White-label delivery

Output in your schema, your identifiers, ready to resell — we stay invisible.

Bulk public records data, collected as a service

Every data company hits the same wall: the catalog roadmap is longer than the collection team. Each new dataset category means new sources, new formats, new breakage to babysit — and the engineers who should be building product end up running collection ops.

Recordpipe sells that layer as a contracted service. You specify the record types, jurisdictions, and schema; we build and operate the pipeline; you own the output. The deliverable is bulk public records data in production shape — corporate filings, UCC filing data, court and judgment records, property records, licensing data — not a scraper handed over for your team to maintain. Because it is contract work rather than a licensed product, the dataset is defined by your catalog's needs instead of whatever a vendor already happens to sell, and the collection headache moves permanently off your org chart.

Normalization is the actual product

Anyone can collect records. The hard, valuable work is making records from hundreds of inconsistent sources behave like one dataset: field mapping into a single schema, date and name standardization, deduplication across sources that publish overlapping records, and stable identifiers that survive refreshes so downstream joins don't shatter every cycle.

That discipline is where Recordpipe earns the contract. Deliverables ship with a documented data dictionary, per-record source jurisdiction and capture timestamps, and explicit handling rules for the messy cases — amended filings, republished records, source-side corrections. Your QA team gets acceptance criteria written into the contract, not a shrug. The result is data your catalog can ingest without a cleanup project on your side, in your field names and your identifier scheme, indistinguishable from data your own team built — because contractually, it is yours.

Refresh operations and delta delivery

A corpus is a photograph; a catalog is a subscription. The recurring value in records data is the refresh — and refresh operations are the grind data companies most want off their plate, because sources change constantly and silently.

Recordpipe runs refresh as managed operations on a contracted cadence, daily or weekly per source. Deliveries arrive as clean deltas — new records, changed records with the changed fields identified, and records no longer published — so your ingestion applies a diff instead of reprocessing a full corpus. Source monitoring is part of the service: when a source changes how it publishes, fixing collection is our job under the contract, not a ticket in your backlog. The same pipeline discipline runs our own products at over one million public records processed nightly, so a contracted refresh program rides infrastructure that is already exercised in production every night.

White-label programs, gap-filling, and category launches

Contract collection fits the moves data companies actually make. Gap-filling: the jurisdictions and record types your current suppliers can't or won't cover, delivered in the schema your catalog already uses. Category launch: a dataset you've wanted to sell — UCC filing data, judgments, licensing records — built as a complete corpus plus ongoing refresh, letting you announce a product without standing up a collection team first. And white-label supply: Recordpipe operates invisibly under NDA while the data ships under your brand, in your identifiers.

One boundary applies across every program. Owner names and mailing addresses are public record and are delivered as filed; personal emails are not promised — contact points come through as published in public filings, we ship no scraped private contact data, and enrichment is your call, under your compliance obligations. Programs run from $5,000 builds to $3 million multi-source operations.

What a delivery looks like

FieldDescription
record_idStable Recordpipe identifier that persists across refreshes, plus your catalog's own ID scheme if supplied
record_typeNormalized record category (formation, UCC filing, judgment, deed, license) per the contract taxonomy
jurisdictionSource jurisdiction — state, county, or registry — for every record
names_as_filedEntity and party names exactly as published, plus normalized variants for matching
address_as_filedMailing and situs addresses as published in the public record, standardized for geocoding
filing_dateFiling or recording date as published by the source
delta_statusNew, changed (with changed fields listed), or no longer published — per refresh cycle
source_capture_tsTimestamp of collection for every record — provenance you can pass to your own customers
qa_flagsRecords failing contract acceptance rules, flagged rather than silently dropped
How teams use it

In the field.

Gap-filling at a property-data company

A property-data company's supply team can hand Recordpipe the counties its current suppliers don't cover, receive the missing records normalized into its existing catalog schema, and put the whole footprint on one refresh cadence — closing coverage gaps without adding a single collection engineer.

Category launch at a business-data platform

A business-data platform that wants to sell UCC filing data can contract the full build: historical corpus, entity matching against its existing company records, and weekly delta refresh. The platform announces the new product line while Recordpipe operates the supply chain behind it.

Refresh outsourcing at a legal-data provider

A legal-data provider maintaining judgment and lien datasets in-house can move the refresh operation to contract: same schema, same cadence, deltas into the existing ingestion path. The internal team stops firefighting source breakage and goes back to building the product customers pay for.

White-label supply for a niche vertical product

A vertical SaaS company can embed a public-records dataset in its product without becoming a data company: Recordpipe builds and refreshes the corpus under NDA, delivery lands in the product's own identifiers, and the customer never sees a supplier's name anywhere.

Data licensing, data reseller program, data partnership: bulk public records without building collection

Our own products process over one million public records nightly. Contract collection runs on the same pipelines, the same normalization discipline, and the same monitoring.

Do you resell the data you collect for us?
Custom deliverables are contracted work product, and exclusivity, ownership, and resale rights are set explicitly per contract — negotiated up front, not discovered later.
What's the largest program you'd take?
Contracts run to $3 million — multi-year, multi-source collection and refresh programs are the top of our published range, and the scoping process is the same regardless of size.
What happens when a source changes and collection breaks?
That's ours under the contract. Source monitoring and repair are part of refresh operations — you see it, if at all, as a note in the delivery log, not as a gap in your catalog.
What formats do you deliver in?
CSV, Parquet, JSON, or database-ready loads — in your schema and identifier scheme. White-label programs deliver in exactly the format your catalog already publishes.
How do deltas work?
Each refresh cycle delivers new records, changed records with the changed fields identified, and records the source no longer publishes — so your ingestion applies a diff rather than reprocessing everything.
Do deliveries include contact data?
Owner names and mailing addresses are public record and are delivered as filed. Personal emails are not promised — contact points appear as published in public filings, we ship no scraped private contact data, and enrichment is your call.
How do we evaluate quality before committing?
The $500 scoping returns a sample in your target schema within 5 business days, and acceptance criteria are written into the contract so QA is a defined gate, not a debate.