Web scraping in R

Web scraping in R is an optional course in the Data science for business postgraduate programme at the Faculty of Economic Sciences of the University of Warsaw. The course covers the basics of scraping in R, including navigation via XPath, handling robots.txt, the rvest and RSelenium packages (and when to use which), working with APIs and associated data formats (JSON), good practices and ethical conduct, as well as some additional, handy packages (e.g. RCrawler, robotstxt). The course is updated each year accordingly with new developments in R and scraping.

Wojciech Hardy has been conducting the classes since the first edition of the programme in 2017/18. 

Scroll to Top