Does anyone know whether I can scrape this site or this one with httr and rvest, or should I use selenium or phantomjs?
Both of the sites seem to be using ajax, and I cant seem to get through it.
Essentially what I am after is the following:
# I want this to return the titles of the listings, but I get character(0)
"https://www.sahibinden.com/satilik" %>%
read_html() %>%
html_nodes(".searchResultsItem .classifiedTitle") %>%
html_text()
# I want this to return the prices of the listings, but I get 503
"https://www.hurriyetemlak.com/konut" %>%
read_html() %>%
html_nodes(".listing-item .list-view-price") %>%
html_text()
Any ideas with v8, or artificial sessions are welcome.
Also, any purely curl solutions are also welcome. I'll try to translate them into httr later :)
Thanks