Skip to content

Repository files navigation

Google Podcast Crawler

Scrapers

Selenium extract the HTML and Symfony Dom Crawler extract data.

  1. Extract podcasts using main page
  2. Extract episodes using podcast page
  3. Extract episode data using episode page

Commands

  • Start Selenium: java -jar selenium-server-standalone-3.141.59.jar
  • Start Elasticsearch: ~/elasticsearch-7.8.0/bin/elasticsearch

About

No description, website, or topics provided.

Resources

Stars

0 stars

Watchers

1 watching

Forks

Releases

Packages

Contributors

Languages