Web Scraping Using R

Web Scraping Using R
复制标题

DOI:
10.1177/2515215919859535
复制
发表时间:
2019-09-01
影响因子:
13.6
通讯作者:
James, Richard J. E.
James, Richard J. E.
中科院分区:
心理学1区
文献类型:
--
作者:
Bradley, Alex;James, Richard J. E.

文献摘要

被引文献

相似文献

互联网在日常生活中无处不在的使用意味着现在有大量的数据可以为人类行为提供新的见解。阻碍更多研究人员利用在线数据的主要障碍之一是他们不具备访问数据的技能。本教程通过提供使用流行的统计语言r抓取在线数据的实用指南来解决这一问题。Web抓取是自动从网站收集信息的过程。这些信息可以采用数字、文本、图像或视频的形式。本教程向读者展示了如何下载网页,从这些页面中提取信息,存储提取的信息,以及在网站的多个页面上这样做。我们建立了一个网站,帮助读者学习如何抓取网页。本网站包含一系列示例,说明如何抓取单个网页和如何抓取多个网页。这些例子附有描述所涉及过程的视频和练习,以帮助读者增加他们的知识和练习他们的技能。示例R脚本已经在开放科学框架中提供。
The ubiquitous use of the Internet in daily life means that there are now large reservoirs of data that can provide fresh insights into human behavior. One of the key barriers preventing more researchers from utilizing online data is that they do not have the skills to access the data. This Tutorial addresses this gap by providing a practical guide to scraping online data using the popular statistical language R. Web scraping is the process of automatically collecting information from websites. Such information can take the form of numbers, text, images, or videos. This Tutorial shows readers how to download web pages, extract information from those pages, store the extracted information, and do so across multiple pages of a website. A website has been created to assist readers in learning how to web-scrape. This website contains a series of examples that illustrate how to scrape a single web page and how to scrape multiple web pages. The examples are accompanied by videos describing the processes involved and by exercises to help readers increase their knowledge and practice their skills. Example R scripts have been made available at the Open Science Framework.