Public sources only
We collect public pages and public records. Nothing is collected from behind a login, a form or a paywall.
We collect only what companies publish themselves, from public sources, with a crawler that identifies itself and obeys robots.txt on every request.
We collect public pages and public records. Nothing is collected from behind a login, a form or a paywall.
Every record comes from the company itself or from a public filing. We do not buy or collect data from brokers, platforms, marketplaces or social networks.
The datasets describe companies, not people. Contact details are removed at collection, and individuals named in filings are not read.
Every input is public at the time it is observed. We receive nothing from insiders, from companies under confidentiality or from any non-public source.
Collection rules are enforced in code, in a single component that every request passes through.
User-agent: FokalsBot rule to your robots.txt. Instructions are on the FokalsBot page.No. All inputs are public at the time they are observed. We receive no information from insiders, from companies under confidentiality or from any non-public source.
No personal data is delivered. Contact details in job descriptions are removed before storage, and the individuals named in SEC filings are not read. A leadership change is recorded as the company’s dated statement and the role concerned.
robots.txt governs every request on every path. Where a site’s network protection refuses a direct request, the page may be retrieved through a proxy or a third-party fetching service, under the same crawler identity and the same robots.txt rules. We do not solve challenges, and a site that refuses on every path is not collected.
From the public actions of companies: website changes, job postings, funding filings and their own announcements. They are not derived from browsing behaviour, bidstream data, cookies or any consumer data.
An evaluation model answers typed questions about a page or a posting under a named, frozen version that is verified on every run. Accuracy is measured by hand for each version. Model output is identified as such in the data.
The sourcing and compliance statement, the methodology and the data dictionary are public. We complete due diligence questionnaires on request.