Supporting Research Data Collection from YouTube with TubeKit
Bibliographic Data
| ID | 12972115 |
|---|---|
| Authors | Chirag Shah (0000-0002-3797-4293, Rutgers, the State University of New Jersey, corresponding author) |
| Year | 2010 |
| Volume | 7 |
| Issue | 2-3 |
| Pages | 226-240 |
| Publication date | 2010-05-18 |
| Peer Reviewed | Yes |
| Open Access | Yes |
| Type | ARTICLE |
| Venue | Journal of Information Technology & Politics (JOURNAL) |
| Journal identifiers | ISSN: 1933-169X • E-ISSN: 1933-1681 |
| Publisher | Routledge (PUBLISHER • GB) |
| DOI | 10.1080/19331681003748875 |
| OpenAlex | W1996272944 |
| Language | EN |
| Citations received | 5 |
We present TubeKit, a query-based YouTube crawling toolkit. This software is a collection of tools that allows users to build their own crawler that can crawl YouTube based on a set of seed queries and collect up to 17 different attributes. TubeKit assists in all the phases of this process, starting with database creation to finally giving access to the collected data with browsing and searching interfaces. We further demonstrate how we used this toolkit to collect elections-related data from YouTube for nearly two years. Some analysis of the collected data relating to the elections is also given
Crawling · Data collection · Data science · Information retrieval · Process (computing · Set (abstract data type · Web crawler · World Wide Web · Advanced Malware Detection Techniques · Computer Science · Video Analysis and Summarization · Web Data Mining and Analysis · Software
| Unique citing works | 5 |
|---|---|
| Citations per year | 0,36 |
| Citation span | 2012 - 2021 (10) |
| Citation velocity | historical |
| Highly cited | No |
| Citation types | Neutral: 5 |