About the Gambling Data Finder
What is this tool?
The Gambling Data Finder is a research tool for the UKRI-funded Youth Gambling Harm Research and Innovation Partnership (Y-GHRIP). It helps researchers, policy teams, practitioners and data owners find, assess and understand datasets that measure gambling participation or gambling-related harm. Coverage runs across the full age range: school surveys, birth cohorts and youth studies alongside adult population surveys, household panels and linked administrative records.
The site is built around metadata: information about studies, questionnaires, variables, measures, documentation and access routes. It does not host individual-level data. Instead, it helps users see what data exist, how the data were collected, what gambling-related content is available, and where to go next if they want to request access.
The resource was developed from Y-GHRIP's data availability and evidence review work. It turns that mapping work into a searchable catalogue that can support evidence reviews, research proposals, data access requests, measure development and future collaboration with study teams.
Purpose
The purpose of the Gambling Data Finder is to make existing gambling-related data easier to discover and compare. It brings together metadata from cohort studies, panel studies, population surveys and other relevant sources so users can identify datasets that may be suitable for research on gambling participation, gambling-related harm, risk and protective factors, and participant characteristics.
The tool supports Y-GHRIP by highlighting what is already available, where there are gaps in measurement or coverage, and where further engagement with data owners may improve the research infrastructure for prevention, policy and practice.
What you can find
- An overview of each dataset or study, including population, geography, design and age coverage
- Relevant survey questions and variables, filterable by study, topic, measure, age group and access route
- Questionnaire wording, response options and documentation links where available
- Information on study waves, data collection years and participant characteristics
- Details of standardised gambling measures and recurring instruments, including PGSI, DSM-derived measures and SOGS-RA
- Risk and protective factor variables relevant to gambling harm
- Repository, licence, governance and contact information to support data access requests
- A basket for saving variables, questions and questionnaires of interest for later review or export
How the metadata helps
By bringing study documentation into one searchable structure, the site helps users judge whether a dataset may be useful before investing time in a full access request. Users can compare how gambling was measured, where questions appeared, whether standardised scales were used, and which contextual variables are available alongside gambling measures.
The catalogue also improves transparency by recording the source documentation used for each entry and by distinguishing extracted metadata from areas that still require manual review or data-owner confirmation.
What it does not do
- It does not host restricted individual-level data unless explicitly allowed
- It does not replace repository or study-team access processes
- It does not guarantee that every gambling-relevant item has been identified
- It does not determine whether a study is suitable for a specific analysis without further review of documentation and access conditions
- It does not provide derivation code unless supplied by the data owner
How to use the site
- Search
- Find variables and questions by keyword across all indexed studies. Search matches variable names, labels, measure names, study names and categories.
- Explore
- Browse gambling topics and risk/protective factor domains without needing to know specific keywords.
- Visualisations
- Review coverage, access, geography, age groups and search-result summaries to identify strengths and gaps in the current evidence base.
- Measures
- View standardised gambling instruments and recurring measures, including PGSI, DSM-derived measures and SOGS-RA.
- Coverage
- See what metadata is available, what has been extracted, and where manual review or further data-owner engagement may be needed.
- Analysis
- Check how ready each dataset is for secondary analysis, scored on measurement detail, access route, risk/protective factor coverage and published use, and see which datasets look compatible enough to pool or compare.
- Ideas
- Turn the catalogue into candidate research. A gap matrix shows which risk and protective factor domains are well covered and which are barely studied, and a question generator suggests cross-study questions the indexed data could answer.
- Basket
- Collect variables, questions and questionnaires of interest and download the selection to support data access requests or evidence mapping.
Data sources
Variable metadata was extracted from questionnaires, codebooks, data dictionaries and other documentation published by study teams. Access and governance information was gathered from study websites, data repositories and open data archives including UK Data Service, ICPSR and related catalogues.
The catalogue is a living resource for Y-GHRIP evidence mapping and future phases of the partnership. If you notice errors or missing data, or would like your study included, please get in touch.
Contact
For questions, corrections or contributions, please contact the Y-GHRIP data availability team at cyp.ghrip@ed.ac.uk.