<?xml version="1.0"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.2 20190208//EN" "JATS-journalpublishing1.dtd" [
]>
<article xml:lang="en" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML"
  dtd-version="1.2" article-type="abstract">
  <front>
    <journal-meta>
      <journal-id journal-id-type="publisher-id">IJPDS</journal-id>
      <journal-title-group>
        <journal-title>International Journal of Population Data Science</journal-title>
        <abbrev-journal-title>IJPDS</abbrev-journal-title>
      </journal-title-group>
      <issn pub-type="epub">2399-4908</issn>
      <publisher>
        <publisher-name>Swansea University</publisher-name>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.23889/ijpds.v9i5.2600</article-id>
      <article-id pub-id-type="publisher-id">9:5:116</article-id>
      <title-group>
        <article-title>Measuring error in the Demographic Index – steps towards a future population statistics system in England and Wales</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <name>
            <surname>Archer</surname>
            <given-names initials="R">Rosalind</given-names>
          </name>
          <xref ref-type="aff" rid="affil-1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <name>
            <surname>Diyasena</surname>
            <given-names initials="P">Peshali</given-names>
          </name>
          <xref ref-type="aff" rid="affil-1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <name>
            <surname>Grant</surname>
            <given-names initials="E">Emma</given-names>
          </name>
          <xref ref-type="aff" rid="affil-1">1</xref>
        </contrib>
      </contrib-group>
      <aff id="affil-1"><label>1</label><institution>Office for National Statistics</institution></aff>
      <pub-date date-type="pub" publication-format="electronic">
        <day>18</day>
        <month>09</month>
        <year>2024</year>
      </pub-date>
      <pub-date date-type="collection" publication-format="electronic">
        <year>2024</year>
      </pub-date>
      <volume>9</volume>
      <issue>5</issue>
      <elocation-id>2600</elocation-id>
      <permissions>
        <license license-type="open-access" xlink:href="https://creativecommons.org/licences/by/4.0/">
          <license-p>This work is licenced under a Creative Commons Attribution 4.0 International License.</license-p>
        </license>
      </permissions>
      <self-uri xlink:href="https://ijpds.org/article/view/2600">This article is available from the IJPDS website at: https://ijpds.org/article/view/2600</self-uri>
    </article-meta>
  </front>
  <body>
    <sec>
      <title>Objective</title>
      <p>We are developing methods to measure the quality of the Demographic Index at a National Statistics Institute. The Demographic Index is the foundational dataset for a future, transformed, population statistics system for England and Wales. It clusters records for individuals across many years of health, education, and tax data, and assigns them a unique ID.</p>
      <p>Previous research has helped us to conceptualise three types of error in the Demographic Index: clustering error, coverage error, and data measurement error. We are currently working on estimating one type of clustering error, whereby records for two individuals are mistakenly assigned the same ID (“false positive clusters”).</p>
    </sec>
    <sec>
      <title>Approach</title>
      <p>Previous work has allowed us to identify variables associated with false positive clusters. We are now working on a stratification method, whereby each ID in the Demographic Index is scored according to how likely it is to contain this error. We plan to measure uncertainty using bootstrapping.</p>
    </sec>
    <sec>
      <title>Results and Conclusions</title>
      <p>This work is ongoing, and we hope to obtain results by summer 2024.</p>
    </sec>
    <sec>
      <title>Conclusion</title>
      <p>It is vital for us to understand and measure error in the Demographic Index because this is necessary for measuring error in any statistics derived from the Index. And it is crucial that population statistics made by any future system should have measures of error. Assuming that our work shows promise, we will continue by measuring other types of error, and by creating use cases to test how these measures can be fed forward into secondary analyses (e.g. population estimation).</p>
    </sec>
  </body>
</article>