Published September 7, 2026 | Version v350

Ipseity Daily Data

Authors/Creators

  • 1. Stony Brook University

Description

Ipseity Daily data

Daily surveys ask American adults whether identity signifiers describe them today. Data are released under CC-BY-4.0. All files are UTF-8 CSV compressed with gzip; missing values are NA.

Canonical microdata: ipseity.csv.gz

One row is one Yes/No response to a signifier on an observation date. The unique key is hashed_respondent_id, observation_date, signifier. endorsed is 1 for Yes and 0 for No; skipped answers are absent. observation_date is the Qualtrics start date (YYYY-MM-DD). hashed_respondent_id is the first 12 hexadecimal characters of a SHA-256 hash of the participant identifier. Raw participant and study identifiers are excluded.

demographics_status is available, consent_revoked, or data_expired. Unavailable demographic values are NA; valid responses still contribute to overall results. age is numeric age in years; sex, ethnicity, birth_country, residence_country, nationality, language, student, and employment are Prolific demographic fields. time_on_task and approvals are Prolific export fields for time taken and total approvals, respectively; these are retained as supplied and may be missing.

Responses require completed surveys and consent to continue. If any respondent/date/signifier key is duplicated, the entire respondent-day is excluded. Demographics are joined by respondent and observation date; current collection retrieves the exact Prolific studies identified in Qualtrics.

Aggregates: ipseity-monthly.csv.gz and ipseity-all-time.csv.gz

One row per signifier and calendar month, or per signifier over all dates. month_start (monthly only) is the first calendar day. yes_count and no_count count responses; observation_count is their sum. respondent_count counts unique respondents and respondent_day_count counts unique respondent/date pairs within the group. prevalence_per_10k is 10000 times yes_count divided by observation_count, without display rounding. earliest_observation_date and latest_observation_date describe actual coverage. Respondent counts must not be summed across groups to estimate unique people.

Months without observations are absent, not zero prevalence. The latest month may be incomplete. Overall statistics include responses with missing demographics; demographic comparisons use only the relevant available values and report their sample sizes. There is a documented collection outage from 2026-08-31 through 2026-09-03.

Observation coverage: 2025-07-08 through 2026-09-06. Version prepared 2026-09-07.

Files

Files (5.0 MB)

Name Size Download all
md5:f118db65287f6a2428c8bf2ae1a01f88
19.0 kB Download
md5:1b2ebaa052cbde0b33642bcc7cc9d6ed
153.7 kB Download
md5:bf2a4c2453c449cf5ccd9dd6bb155df5
4.9 MB Download

Additional details