Sorting URLs out
Seeing the web through infrastructural inversion of archival crawling
Bibliographic Data
| ID | 13119005 |
|---|---|
| Authors | Emily Maemura (0000-0002-9329-7995, University of Illinois Urbana-Champaign, corresponding author) |
| Year | 2023 |
| Volume | 7 |
| Issue | 4 |
| Pages | 386-401 |
| Publication date | 2023-09-16 |
| Peer Reviewed | Yes |
| Open Access | No |
| Type | ARTICLE |
| Venue | Internet Histories (JOURNAL) |
| Journal identifiers | ISSN: 2470-1475 • E-ISSN: 2470-1483 |
| Publisher | Routledge (PUBLISHER • GB) |
| DOI | 10.1080/24701475.2023.2258697 |
| OpenAlex | W4386802863 |
| Language | EN |
| Citations received | 4 |
| References cited | 37 |
Web archives collections have become important sources for Internet scholars by documenting the past versions of web resources. Understanding how these collections are created and curated is of increasing concern and recent web archives scholarship has studied how the artefacts stored in archives represent specific curatorial choices and collecting practices. This paper takes a novel approach in studying web archiving practice, by focusing on the challenges encountered in archival web crawling and what they reveal about the web itself. Inspired by foundational work in infrastructure studies, infrastructural inversion is applied to study how crawler interactions surface otherwise invisible, background or taken-for-granted aspects of the web. This framework is applied to study three examples selected from interviews and ethnographic fieldwork observations of web archiving practices at the Danish Royal Library, with findings demonstrating how the challenges of archival crawling illuminate the web's varied actors, as well as their changing relationships, power differentials and politics. Ultimately, analysis through infrastructural inversion reveals how collection via crawling positions archives as active participants in web infrastructure, both shaping and shaped by the needs and motivations of other web actors
Algorithm · Biology · Crawling · Sorting · Web crawler · World Wide Web · Computer Science · Digital and Cyber Forensics · Digital and Traditional Archives Management · Web Data Mining and Analysis · Anatomy
Media Technologies
Steps Toward an Ecology of Infrastructure
The Oxford Handbook of Internet Studies
Sorting Things Out
All Warc and no playback
Unearthing the Infrastructure
Little history of CAPTCHA
The invention of the archived web
The Sage handbook of web history
The Internet Archive and the socio-technical construction of historical facts
Everything on the internet can be saved”
Norm conflict in the governance of transnational and distributed infrastructures
Radical infrastructure
Infrastructure studies meet platform studies in the age of Google and Facebook
The Politics of Mass Digitization
How cells became records
Neutrality, social justice and the obligations of archival education and educators in the twenty-first century
Archival assemblages
Data Labours
Institutional Ecology, `Translations' and Boundary Objects
Web 25
| Unique citing works | 4 |
|---|---|
| Citations per year | 2 |
| Citation span | 2024 - 2026 (3) |
| Citation velocity | current |
| Highly cited | No |
| Citation types | Neutral: 4 |