Data for academic research, documented well enough to build a study on.
“The only dataset with the depth and breadth for powering quantitative analysis.” — USC Marshall
Data for academic research has to hold up to a level of scrutiny commercial buyers rarely apply. Where did it come from, how was it collected, what are the actual gaps. We publish fill rates by field and refresh hourly, so a researcher can evaluate the dataset the way they would evaluate any other data source before building a study on it.
· Live sample of the production index
Representative of the production index, not a real customer.


Fill rates by field, not one blended accuracy claim.
A blanket accuracy number tells a researcher nothing about the field they actually need. We publish coverage field by field, on every dataset page, including the fields where it is thin.
Where each field came from
Every field is timestamped and refreshed hourly, so a dataset pulled today is dated accurately rather than presented as static.
See the full People dataset →Emails, direct dials, and posts are priced separately from the standard subscription.
People Data, field by field. Every dataset publishes its own.
Public data, documented by field.
We build four datasets from public information, relevant to research use: people data, company data, job postings, and mobile app and SDK data. Each field carries a published fill rate rather than a blanket accuracy claim, so a researcher can see exactly what's covered and what isn't before designing a study around it.
Real research, not a hypothetical use case
People and company data come from public, voluntarily shared information, never gated or login-required data. Job postings are tracked as they're published. Mobile app and SDK data comes from decompiling apps directly, verified by hand, not inferred from a pattern that happens to correlate. A small number of specific fields, like email and phone, come through licensed partnerships to close gaps we can't otherwise close. All of it is timestamped and refreshed hourly, so a dataset pulled today is dated accurately, not presented as static.
Job postings and people data for labor market research
Job postings data includes salary information where the original posting disclosed it, along with location, seniority, and full description text, useful for labor-market research beyond just counting open roles. People data adds employment history, role tenure, and location, so a study can track workforce movement over time, not just a single snapshot.
An SDK dataset built from decompiled apps, not store metadata
Our mobile app and SDK data comes from decompiling apps directly, an app store metadata dataset built on what's actually inside an app, not just what's listed in a store description. This is the same dataset currently being used in an active academic research project, not a separate research-only product.
How a public signal becomes a dataset you can cite.
Collect public data
crawl · decompile · license
Timestamp every field
refreshed hourly
Publish fill rates
by field, not blended
Deliver to the study
API · flat file · hosted
Built on four datasets.
People Data
1.20B+ professional profiles, refreshed hourly.
Talk to us about People Data →Company Data
120M+ companies tracked (40M+ active), refreshed hourly.
Talk to us about Company Data →Job Postings
1.8B+ postings tracked, 24.5M+ open now, refreshed hourly.
Talk to us about Job Postings →Mobile Apps & SDKs
24M+ apps decompiled, 100K+ SDKs hand-cataloged.
Talk to us about Mobile Apps & SDKs →Where this shows up in real research.
- Individual Researcher Licensing. Professors and PhD researchers using people, company, or SDK data to power their own studies.
- Labor Market & Workforce Research. Job postings and people data for research on employment trends and workforce movement.
- University & Institutional Licensing. Job-postings data for career-services and workforce-insight use, a distinct conversation from individual research access.
- Alumni & Community Search. An emerging pattern, semantic search over an institution's own population using people data.
Four datasets, documented field by field.
Job postings go back to 2021, mobile app and SDK data to 2008, company data to 2013, and people data to 2013. All of it is timestamped and refreshed hourly, so a dataset pulled today is dated accurately rather than presented as static.
Evaluate the data before you build the study.
Request a research sample, or book a short call to talk through a specific project. Bring a research question or a dataset you're evaluating, and we'll show you what's actually there.
no credit card · founder-owned · we read every reply