Scraping IMDB With Python 2024. without selenium!

Pubblicato il: 03 giugno 2024
sul canale di: CodeMate TV
3,075
71

In this video, we'll be diving into the world of web scraping by extracting IMDb data using Python Scrapy and uncovering hidden information on websites! 🚀 No need for Selenium, Playwright, or any automated web drivers – we'll be sending requests, cleaning the data, and converting it into JSON format. 🧹✨

If you enjoyed this video and want to support me, buy me a coffee! ☕️
https://codematetv.com/donate

I'll also show you how to bypass website restrictions using a fake user agent! 🕵️‍♂️🔍

🌟 Important: Always consider the ethical implications of web scraping. 🌟

📜 Complete Code: https://github.com/itishosseinian/imd...

📚 Scrapy User Agent Library: https://pypi.org/project/scrapy-user-...

💬 Feel free to contact me if you have any questions about web scraping! 💬

00:00 - Intro
01:53 - Our job and checking the final result
03:50 - Setting up a virtual environment and spider
09:23 - Checking how data is released in IMDB and what we need to scrape
16:20 - Working with Scrapy shell to find the required elements
19:35 - Bypassing restriction such as robots.txt and user-agent
27:00 - Cleaning data using json in Scrapy shell
31:10 - Implementing our spider
40:35 - Working with items.py and pipelines.py to clean and convert seconds to real time format
51:01 - Outro Chat (A Little Talk)


In questa pagina del sito puoi guardare il video online Scraping IMDB With Python 2024. without selenium! della durata di ore minuti seconda in buona qualità , che l'utente ha caricato CodeMate TV 03 giugno 2024, condividi il link con amici e conoscenti, su youtube questo video è già stato visto 3,075 volte e gli è piaciuto 71 spettatori. Buona visione!