MétaCan
Menu
Back to cohort
Record W6967878049 · doi:10.5255/ukda-sn-7403-2

Citizenship Survey, 2005-2011: Secure Access

2023· dataset· en· W6967878049 on OpenAlexaboutno aff

Bibliographic record

VenueUK Data Archive · 2023
Typedataset
Languageen
Field
Topic
Canadian institutionsnot available
Fundersnot available
KeywordsQuarter (Canadian coin)Government (linguistics)CitizenshipWork (physics)Data collectionSample (material)Survey data collectionAmerican Community Survey

Abstract

fetched live from OpenAlex

The Citizenship Survey (known in the field as the Communities Study) ran from 2001 to 2010-2011. It began as the 'Home Office Citizenship Survey' (HOCS) before the responsibility moved to the new Communities and Local Government department (DCLG) in May 2006. The survey provided an evidence base for the work of DCLG, principally on the issues of community cohesion, civic engagement, race and faith, and volunteering. The survey was used extensively for developing policy and for performance measurement. It was also used more widely, by other government departments and external stakeholders to help inform their work around the issues covered in the survey. The survey was conducted on a biennial basis in 2001, 2003, 2005 and 2007-2008. It moved to a continuous design in 2007 which means that data became available on a quarterly basis from April 2007. Quarter one data were collected between April and June; quarter two between July and September; quarter three between October and December and quarter four between January and March. Once collection for the four quarters was completed, a full aggregated dataset was made available, and the larger sample size allowed more detailed analysis. In January 2011, the DCLG announced that the Citizenship Survey was to close. As part of the drive to deliver cost savings across government and to reduce the fiscal deficit, research budgets were closely scrutinised to identify where savings can be made. For this reason, and the belief that priority data from this survey could either be dropped; collected less frequently; or collected via other means, the survey was cancelled. Fieldwork concluded on 31 March 2011, followed by publication of reports in the months after analysis of that data. Further information about the survey, including links to publications, can be found on the National Archives webarchive page for the Citizenship Survey. Detailed topic reports are published for the years up to 2009-2010 and there are statistical releases which cover 2010-2011. The Consultation outcome: the future of the citizenship survey statement can be viewed on the gov.uk website. The Community Life Survey, (held under GN 33475), which began in 2012-2013 and is conducted by the Cabinet Office, incorporates a small number of priority measures from the Citizenship Survey, in order that trends in these issues can continue to be tracked over time. For these measures the Community Life Survey findings are comparable to the Citizenship Survey findings. UK Data Archive holdings: End User Licence and Secure Access The Archive holds standard End User Licence (EUL) versions of the complete Citizenship Survey series from 2001-2011, held under SNs 4754, 5087, 5367, 5739, 6388, 6733 and 7111. Currently, the Secure Access version (SN 7403) includes only the 2005, 2007-2008, 2008-2009 and 2009-2010 and 2010-2011 waves. The Secure Access datasets include extra variables that are not available in the standard EUL versions. The extra variables comprise: more detailed and extensive household and demographic information; more detailed geographies, including Police Force Area, Local Authority Districts, Wards, Health Areas, Middle Layer Super Output Areas (MSOA) and Lower Layer Super Output Areas (LSOA); more detailed responses to questions covering extremism, immigration, and religion; and more detailed administrative variables. Prospective users of the Secure Access version of the Citizenship Survey will need to agree to rigorous Terms and Conditions, including applying for ESRC Accredited Researcher Status and attending a training session, in order to obtain permission to use that version (see 'Administrative and Data Access' section in the catalogue record). Therefore, users are encouraged to download and inspect the EUL versions of the data prior to ordering the Secure Access version. For the second edition (March 2019) data and documentation for survey years, 2005, 2007-2008 and 2008-2009 have been added.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.004
metaresearch head score (Gemma)0.014
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: none
GenreCandidate signal: Dataset · Consensus signal: Dataset
Teacher disagreement score0.070
Threshold uncertainty score0.140

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0040.014
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.001
Bibliometrics0.0030.007
Science and technology studies0.0010.000
Scholarly communication0.0020.002
Open science0.0010.002
Research integrity0.0010.002
Insufficient payload (model declined to judge)0.0220.032

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.169
GPT teacher head0.378
Teacher spread0.209 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designObservational
Domainnot available
GenreDataset

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2023
Admission routes1
Has abstractyes

Explore more

Same venueUK Data ArchiveFrench-language works237,207