I am trying to scrape parts of speeches held in parliament with the rvest package.
Using the css selector or chrome's inspector tool provide me with a selector, however I am unable to retrieve the intended (any) data. AFAIK, the site is also not java etc based, i.e. no RSelenium etc should be required.
here is the link:
library(tidyverse)
library(rvest)
library(xml2)
session_1 <- "https://www.parlament.gv.at/PAKT/VHG/XXVII/NRSITZ/NRSITZ_00001/fnameorig_796482.html"
x <- session_1 %>%
rvest::read_html() %>%
rvest::html_element("wordsection14") %>%
rvest::html_text()
Eventually, I would like to be able to get the text contained in all elements with the class 'wordsection*'.
Would be very grateful for any hint. Many thanks.