BeautifulSoup Tutorial - Web Scraping Resumes in Python

Veröffentlicht am: 13 Juni 2020
auf dem Kanal: Web Scraping with Andy
2,828
42

In this Web Scraping Python Tutorial, we will learn how to scrape resumes using the BeautifulSoup library. BeautifulSoup is an awesome library for parsing HTML and getting exactly the data you need.
Are you interested in scraping resumes from Robert Half, CareerBuilder, Indeed, Ladders, Glassdoor, etc? Do you want to scrape FREE resumes in extremely large quantities?

Introducing PostJobFree.com - An undervalued FREE alternative for the above web sites. I have no affiliation with them. I am not making any money on this.

Here is a list of fasts.
✅ This job board allow us to scrape resumes/CVs in extremely large quantities. Let's say 500k resumes or more!
✅ Vary Job title while scraping the resumes/CVs.
✅ Vary locations while scraping the resumes/CVs.
✅ Total Visits: 320K (May 2020)
✅ Traffic by countries: US (54%), India (14%), Canada (5%), ...
❌ Personal information is hidden. I mean phone number, postal code, etc.
✅ Many resumes still contain LinkedIn links, real emails, etc.
❌ This job board provides resumes/CVs in plain text. So the resumes/CVs lose some formatting metadata.

The most popular libraries used by web scraping developers in python are Beautifulsoup, Scrapy, and Selenium, but every library has its own pros and cons. Scrapy is like a spaceship. Beautifulsoup is like a skateboard. If you are a Python beginner, then it wouldn't be right to start with learning how to 'drive' a spaceship. Start with the skateboard first! So that is the reason why we are using Beautifulsoup. We’ll pair BeautifulSoup with the Requests library to fetch resumes/CVs from the job board via HTTP GET requests. This video shows the exact steps to successfully scrape resumes/CVs. You can apply the learned skills while solving freelance web scraping tasks from Upwork.

📄 Code Download: https://github.com/andrei-volkau/Resu...

⭐️Timestamps⭐️
0:00 - Intro
0:54 - Installing Requests and BeautifulSoup
1:16 - Import libs
1:30 - Defining target URL + web site overview
1:48 - Defining our first Get Request
2:21 - Instantiating a BeautifulSoup Object
3:43 - Extract resume links
4:39 - Extract job title
4:49 - Extract resume text
5:31 - Write into a CSV file
5:42 - Specify a delay
5:57 - Scraped resumes!

Let's connect on Linkedin ►   / andreivolkau  
Find me on Upwork ► https://www.upwork.com/o/profiles/use...

⚡ Please leave a LIKE and SUBSCRIBE for more content! ⚡

⭐ Tags ⭐
web scraping
beautifulsoup
scraping
web scraping python
python beautifulsoup
Python Tutorials
Python web scraper tutorial
Web Scraper Python
octoparse
data scraping
python scrapy
web scraping using python
web crawler python

⭐ Hashtags ⭐
#webscraping #beautifulsoup #pythontutorial


Auf dieser Seite können Sie das Online-Video BeautifulSoup Tutorial - Web Scraping Resumes in Python mit der Dauer stunde minuten sekunde in guter Qualität ansehen, das der Benutzer Web Scraping with Andy 13 Juni 2020 hochgeladen hat, den Link mit Freunden und Bekannten teilen, dieses Video wurde auf Youtube bereits 2,828 Mal angesehen und es wurde von 42 den Zuschauern gefallen. Viel Spaß beim Betrachtenden Zuschauern gefallen!