BeautifulSoup Tutorial - Web Scraping Resumes in Python

Pubblicato il: 13 giugno 2020
sul canale di: Web Scraping with Andy
2,828
42

In this Web Scraping Python Tutorial, we will learn how to scrape resumes using the BeautifulSoup library. BeautifulSoup is an awesome library for parsing HTML and getting exactly the data you need.
Are you interested in scraping resumes from Robert Half, CareerBuilder, Indeed, Ladders, Glassdoor, etc? Do you want to scrape FREE resumes in extremely large quantities?

Introducing PostJobFree.com - An undervalued FREE alternative for the above web sites. I have no affiliation with them. I am not making any money on this.

Here is a list of fasts.
✅ This job board allow us to scrape resumes/CVs in extremely large quantities. Let's say 500k resumes or more!
✅ Vary Job title while scraping the resumes/CVs.
✅ Vary locations while scraping the resumes/CVs.
✅ Total Visits: 320K (May 2020)
✅ Traffic by countries: US (54%), India (14%), Canada (5%), ...
❌ Personal information is hidden. I mean phone number, postal code, etc.
✅ Many resumes still contain LinkedIn links, real emails, etc.
❌ This job board provides resumes/CVs in plain text. So the resumes/CVs lose some formatting metadata.

The most popular libraries used by web scraping developers in python are Beautifulsoup, Scrapy, and Selenium, but every library has its own pros and cons. Scrapy is like a spaceship. Beautifulsoup is like a skateboard. If you are a Python beginner, then it wouldn't be right to start with learning how to 'drive' a spaceship. Start with the skateboard first! So that is the reason why we are using Beautifulsoup. We’ll pair BeautifulSoup with the Requests library to fetch resumes/CVs from the job board via HTTP GET requests. This video shows the exact steps to successfully scrape resumes/CVs. You can apply the learned skills while solving freelance web scraping tasks from Upwork.

📄 Code Download: https://github.com/andrei-volkau/Resu...

⭐️Timestamps⭐️
0:00 - Intro
0:54 - Installing Requests and BeautifulSoup
1:16 - Import libs
1:30 - Defining target URL + web site overview
1:48 - Defining our first Get Request
2:21 - Instantiating a BeautifulSoup Object
3:43 - Extract resume links
4:39 - Extract job title
4:49 - Extract resume text
5:31 - Write into a CSV file
5:42 - Specify a delay
5:57 - Scraped resumes!

Let's connect on Linkedin ►   / andreivolkau  
Find me on Upwork ► https://www.upwork.com/o/profiles/use...

⚡ Please leave a LIKE and SUBSCRIBE for more content! ⚡

⭐ Tags ⭐
web scraping
beautifulsoup
scraping
web scraping python
python beautifulsoup
Python Tutorials
Python web scraper tutorial
Web Scraper Python
octoparse
data scraping
python scrapy
web scraping using python
web crawler python

⭐ Hashtags ⭐
#webscraping #beautifulsoup #pythontutorial


In questa pagina del sito puoi guardare il video online BeautifulSoup Tutorial - Web Scraping Resumes in Python della durata di ore minuti seconda in buona qualità , che l'utente ha caricato Web Scraping with Andy 13 giugno 2020, condividi il link con amici e conoscenti, su youtube questo video è già stato visto 2,828 volte e gli è piaciuto 42 spettatori. Buona visione!