The Regionalization and Aggregation of In‐App Location Data to Maximize Information and Minimize Data Disclosure
Dados Bibliográficos
| ID | 21321663 |
|---|---|
| Autores | Louise Sieg (0009-0002-3130-2100, Department of Geography University College London London UK, autor correspondente), James Cheshire (0000-0003-4552-5989, Department of Geography University College London London UK) |
| Ano | 2025 |
| Volume | 57 |
| Fascículo | 1 |
| Páginas | 27-51 |
| Data de publicação | 2025-01-01 |
| Peer Reviewed | Sim |
| Open Access | Sim |
| Tipo | ARTICLE |
| Periódico | Geographical Analysis (JOURNAL) |
| Identificadores do periódico | ISSN: 0016-7363 • E-ISSN: 1538-4632 |
| Editora | Wiley (PUBLISHER • GB) |
| DOI | 10.1111/gean.12406 |
| OpenAlex | W4399435122 |
| Idioma | EN |
| Referências citadas | 49 |
To minimize the disclosure of personal information, sensitive location data collected by mobile phones is often aggregated to predefined geographic units and presented as counts of devices at a given time. The use of grids or units created by statistical agencies for the dissemination of traditional data sets—such as censuses—are common choices for this aggregation process. However, these can result in large variations in the number of devices encapsulated within each geographic unit, resulting in over‐generalization and a loss of information in some areas. To alleviate this issue, we propose a new method for the aggregation of mobile phone generated location data sets that creates bespoke geometries that maximize the granularity of the data, whilst minimizing the risks of disclosing personal information. The resulting small areas are built on Uber's H3 hexagonal indexing system by attributing activity counts and land‐use features to each cell, then merging cells into geographies containing a predetermined number of data points and respecting the underlying topography and land use. This methodology has applications to widely available data sets and enables bespoke geographical units to be created for different contexts. We compare the generated units to established aggregates from the England and Wales Census and Ordnance Survey. We demonstrate that our outputs are more representative of the original mobile phone data set and minimize data omission caused by low counts. This speaks to the need for a data‐driven and context‐driven regionalization methodology
Computer network · Data aggregator · Internet privacy · Computer Science · Data-Driven Disease Surveillance · Human Mobility and Location-Based Analysis · Privacy, Security, and Data Protection
Map Algebra
Dynamic population mapping using mobile phone data
Community detection in graphs
Unique in the Crowd
Computational Social Science
Escaping from Cities during the Covid-19 Crisis
Does Urban Mobility Have a Daily Routine? Learning from the Aggregate Data of Mobile Networks
Big data and human geography
Thirty Years of Geographical (In)consistency in the British Population Census
Geodesic Discrete Global Grid Systems
Towards the Geographies of the 2001 UK Census of Population
Maintaining Existing Zoning Systems Using Automated Zone-Design Techniques
An Empirical Study of Some Zone-Design Criteria
Ecological Fallacies and the Analysis of Areal Census Data
A Case Study of the Impact of Statistical Disclosure Control on Data Quality in the Individual UK Samples of Anonymised Records
| Velocidade de citação | historical |
|---|---|
| Altamente citado | Não |