Skip to content
Data for Academic Research · Universities & Researchers

Data for academic research, documented well enough to build a study on.

“The only dataset with the depth and breadth for powering quantitative analysis.” — USC Marshall

Data for academic research has to hold up to a level of scrutiny commercial buyers rarely apply. Where did it come from, how was it collected, what are the actual gaps. We publish fill rates by field and refresh hourly, so a researcher can evaluate the dataset the way they would evaluate any other data source before building a study on it.

· Live sample of the production index

dataset card · live indexREC·ACD
DatasetPeople Data
Attributes250+ per record
Fill ratespublished per field
Historyback to 2013
Last refresh38 min ago
Samplefree · no signup
illustrative · pulled livedocumented

Representative of the production index, not a real customer.

Public data onlyFill rates published by fieldRefreshed hourlyFree research sampleSee the dataset
Powering the world's best data teams
AdobeOracleAmazonSpaceXClayAngelListBuyerCaddyRox
Check first

Fill rates by field, not one blended accuracy claim.

A blanket accuracy number tells a researcher nothing about the field they actually need. We publish coverage field by field, on every dataset page, including the fields where it is thin.

Where each field came from

CrawledPeople, company, and job-postings data, from public, voluntarily shared information
DecompiledMobile app and SDK data, read out of the app itself and verified by hand
LicensedA small number of specific fields, like email and phone, through partnerships

Every field is timestamped and refreshed hourly, so a dataset pulled today is dated accurately rather than presented as static.

See the full People dataset →
field fill ratesglobal / US
Profile ID / URL / slug100% / 100%
First / last name92% / 91%
Headline88% / 84%
Locality87% / 67%
Connection count63% / 70%
Company name61% / 67%
Title60% / 66%
Experience history81% / 67%
Education39% / 40%
Industry24% / 36%
Skills11% / 17%
Email (priced separately)23% / 41%

Emails, direct dials, and posts are priced separately from the standard subscription.

People Data, field by field. Every dataset publishes its own.

Why MixRank

Public data, documented by field.

We build four datasets from public information, relevant to research use: people data, company data, job postings, and mobile app and SDK data. Each field carries a published fill rate rather than a blanket accuracy claim, so a researcher can see exactly what's covered and what isn't before designing a study around it.

01

Real research, not a hypothetical use case

People and company data come from public, voluntarily shared information, never gated or login-required data. Job postings are tracked as they're published. Mobile app and SDK data comes from decompiling apps directly, verified by hand, not inferred from a pattern that happens to correlate. A small number of specific fields, like email and phone, come through licensed partnerships to close gaps we can't otherwise close. All of it is timestamped and refreshed hourly, so a dataset pulled today is dated accurately, not presented as static.

02

Job postings and people data for labor market research

Job postings data includes salary information where the original posting disclosed it, along with location, seniority, and full description text, useful for labor-market research beyond just counting open roles. People data adds employment history, role tenure, and location, so a study can track workforce movement over time, not just a single snapshot.

03

An SDK dataset built from decompiled apps, not store metadata

Our mobile app and SDK data comes from decompiling apps directly, an app store metadata dataset built on what's actually inside an app, not just what's listed in a store description. This is the same dataset currently being used in an active academic research project, not a separate research-only product.

How it works

How a public signal becomes a dataset you can cite.

01

Collect public data

crawl · decompile · license

02

Timestamp every field

refreshed hourly

03

Publish fill rates

by field, not blended

04

Deliver to the study

API · flat file · hosted

What you get

Where this shows up in real research.

  • Individual Researcher Licensing. Professors and PhD researchers using people, company, or SDK data to power their own studies.
  • Labor Market & Workforce Research. Job postings and people data for research on employment trends and workforce movement.
  • University & Institutional Licensing. Job-postings data for career-services and workforce-insight use, a distinct conversation from individual research access.
  • Alumni & Community Search. An emerging pattern, semantic search over an institution's own population using people data.
By the numbers

Four datasets, documented field by field.

Job postings go back to 2021, mobile app and SDK data to 2008, company data to 2013, and people data to 2013. All of it is timestamped and refreshed hourly, so a dataset pulled today is dated accurately rather than presented as static.

1.20B+People profiles
120M+Companies tracked
1.8B+Job postings tracked
24M+Apps decompiled
100K+SDKs hand-cataloged
HourlyRefresh, flat price
Test the data

Evaluate the data before you build the study.

Request a research sample, or book a short call to talk through a specific project. Bring a research question or a dataset you're evaluating, and we'll show you what's actually there.

no credit card · founder-owned · we read every reply

FAQ

Frequently asked questions

Where can I get job postings or people data for academic research?
Directly from MixRank, as a feed via API, flat file, or hosted delivery. A free research sample is available before any licensing conversation, so a dataset can be evaluated on its own before committing to it.
How is MixRank data collected and documented?
All of it comes from public, voluntarily shared information, never anything gated or private. We publish fill rates by field rather than one blended accuracy claim, and refresh the full dataset hourly, so what you're working with is documented and current, not a static export.
Can I get a data sample before licensing?
Yes. A free sample lookup on the live index, no signup, resolves a real record against a genuine cut of the production dataset, and a research-specific sample can be requested directly for a closer look at a particular dataset.
Do you offer academic or research pricing?
One flat price per dataset, refresh included, sized to the scope of the project rather than metered by how often you read it. Researchers should reach out directly, since research-scale needs often differ from a commercial licensing agreement.
How do I cite MixRank data in a publication?
Cite MixRank as the data provider, along with the specific dataset and the date it was accessed, consistent with how any commercial data source is typically cited in academic work. Reach out directly if a specific citation format is required for a particular publication.
Is the data GDPR compliant for research use?
We collect only publicly available information. Specific compliance requirements for a research use case are best confirmed directly with our team, since the right answer depends on how the data will be used.
How far back does the data go?
Depends on the dataset: job postings back to 2021, company data back to 2013, people data back to 2013, mobile app and SDK data back to 2008.
What delivery formats are available for researchers?
API, flat file (including formats like JSONL), or direct delivery into a hosted database. These are the same delivery options available to a commercial customer, sized to a research project's actual scope.
APIFlat-filePostgreSQL tablesHosted by MixRankTalk to our data team