Continue → Overview

NameGender Research

Published measurements for name-to-gender inference. Each report keeps coverage separate from accuracy, names the fixture, dates the run and shows weak segments instead of averaging them away.

Published reports

Gender Prediction API Benchmark 2026

A country and writing-system breakdown from one fixed 4,610-row fixture. The report includes zero-coverage segments and distinguishes declined answers from incorrect answers.

Coverage
87.87%
Answered accuracy
96.03%
All rows correct
84.38%

Benchmark methodology and holdout test

Definitions, sampling rules, commands and limitations. The separate holdout run uses names excluded from indexed evidence and reports its low coverage alongside answered accuracy.

Holdout names
2,000
Answered accuracy
93%

What every report must disclose

The denominator

Coverage uses all submitted rows. Answered accuracy uses only rows with a prediction. End-to-end accuracy returns to all submitted rows.

The population

Country and writing-system results remain visible because a large Latin-script segment can conceal a serious gap elsewhere.

The test date

Data changes. Reports carry the actual generation date, and sitemap metadata changes only when the measurement does.

The failures

Unknown and wrong answers are part of the result. They are not removed to make the headline number more flattering.

Country-sensitive name profiles

The curated profile collection turns country-level evidence into citation-ready pages for names such as Andrea, Jean, Kim, Deniz and Ayşe. Each page exposes sample size and source beside the classification.

Explore name gender profiles →