Compliance & Data ProvenanceHow we source, license and document public data
Every dataset we deliver comes with a record of where it came from and how it was collected. This page states our position in full, before you have to ask for it.
// WHAT WE COLLECT
PUBLICLY ACCESSIBLE SOURCES ONLY
We collect data that is publicly accessible and manifestly made public by the data subject or the hosting platform. We do not collect from sources that require an account, a paywall, or any form of authorisation we were not granted.
BUSINESS AND PROFESSIONAL DATA
Our collection is limited to publicly available professional and business-related information — firmographics, public professional histories, public listings, filings and pricing. We are a B2B data provider and scope our collection accordingly.
// HOW WE COLLECT IT
PROTOCOLS HONOURED
We honour standard web scraping protocols, including robots.txt where applicable, and we respect per-source rate limits. Collection is designed not to degrade the source.
NO CIRCUMVENTION
We do not defeat access controls, and we do not sell the ability to do so. If a source cannot be collected within its own stated limits, our answer is that we will not collect it.
DOCUMENTED PROVENANCE
Each dataset carries its source category, collection method and refresh cadence, so your own compliance review has something concrete to assess rather than a vendor assurance.
// LAWFUL BASIS & PRIVACY RIGHTS
GDPR — LEGITIMATE INTEREST
For individuals in the EEA and UK, we process publicly available personal data under legitimate interest, Article 6(1)(f). That interest is providing B2B intelligence and data services, balanced against individual privacy rights by restricting collection to publicly available professional data.
CCPA & DATA SUBJECT REQUESTS
Depending on jurisdiction — including the EEA, UK and California — you may have rights of access, correction, deletion and objection. Our Privacy Policy sets out those rights and how to exercise them.
// LICENSING & ACCEPTABLE USE
WHAT YOU LICENSE
We claim no ownership over raw public facts. What we license is the compilation — the schema, the structure and the engineering around it. Terms are set out in the Terms of Service.
HOW IT MAY NOT BE USED
Our data may not be used for unsolicited spam, phishing or harassment, to re-identify anonymised data, to circumvent privacy protections, or to train generative models in violation of a source’s copyright.
Procurement questions
If your legal or security review needs a provenance record for a specific dataset, a Data Processing Addendum, or answers to a vendor questionnaire, ask us directly and we will work through it with you.
