You will create a very simple web scraper to scrape Behance.net
Budget: $10 – $30 USD
You must live in certain countries in the Western Balkans, Eastern Europe, or Latin America.
You must live in...
Albania,
Bosnia and Herzegovina,
Brasil,
Bulgaria,
Chile,
Colombia,
Kosovo,
Mexico,
Moldova,
Montenegro,
North Macedonia,
Romania,
Serbia,
Uruguay, or
Venezuela.
✻ OVERVIEW ✻
You will provide me with a very simple web scraper which I will use to scrape user names from Behance.net.
You may use whatever web scraper you prefer except for Selenium. You must *not* use Selenium.
I will run the web scraper on Lubuntu 20.04 LTS on my laptop computer.
The web scraper will append user names to a text file named, "behance_user_names.txt".
For example, behance_user_names.txt will contain data as follows...
https://www.behance.net/svitlnart
https://www.behance.net/demsey_dee
https://www.behance.net/katebel
and so on.
✻ DETAILS ✻
For example, please go to...
https://www.behance.net/search/projects?country=UA&field=illustration
Currently I obtain user names by right clicking on a user name and then choosing "Copy link address". Please see... https://i.imgur.com/8QIKnSX.png
Then I go to a URL such as this... https://www.behance.net/svitlnart so that I can review that illustrator's portfolio.
The phrase "plusieurs propriétaires is French for "multiple owners" in English. Please see... https://i.imgur.com/6mowuJ5.jpg. The web scraper will ignore URLs when a project has more than one owner (plusieurs propriétaires or multiple owners). In other words, I don't need URLs when a project has more than one owner (plusieurs propriétaires or multiple owners). But I need URLs when a project has one owner.
The web scraper must scrape all of the URLs on a page for projects that have one owner. What do I mention this? Isn't this obvious? I want to ensure you understand this requirement. Behance.net uses infinite scroll https://en.wiktionary.org/wiki/infinite_scroll. Therefore, the web scraper must continue to scroll down each page and scrape user names until the end of the page. I want all of the "infinite" results on the page I am scrapping. Obviously, the results aren't infinite; the results are obviously finite.
✻ SCREENCAST ✻
In addition to the web scraper, you will provide me with a screencast that is between one minute and ten minutes.
The screencast does *not* need audio. (You don't need to narrate the screencast). However, if you prefer, you may narrate the screencast. Are you confused? If you don't want to speak English then you don't need to speak English. See, I realize many programmers are uncomfortable speaking English.
The screencast will clearly explain to me how to use the web scraper.
The screencast will also show me how to install all necessary applications necessary to use the web scraper on Linux. (I use Lubuntu 20.04 LTS).
The screencast must *not* be created on Windows or Mac. The screencast must be created on Linux.
The screencast must not be larger than 50 megabytes yet the screencast must be easy for a person with normal vision to understand. In other words, you must not compress the screencast so much that it is difficult for a person with normal vision to understand.
✻ HOW TO APPLY ✻
If you are interested in working on this project with me then please copy and paste...
*** I am interested in this project to scrape data from Behance.net. Due to a glitch with Freelancer, I might not receive an email notification from Freelancer indicating that you have responded to me. Therefore, I might need to login to Freelancer to find out that you have responded to me.***
...into a message you send me. (Please do not send me the three asterisks before and after the text).
Please do *not* bother wasting your time including any other information. Instead, simply copy and paste the information above.
You must live in...
Albania,
Bosnia and Herzegovina,
Brasil,
Bulgaria,
Chile,
Colombia,
Kosovo,
Mexico,
Moldova,
Montenegro,
North Macedonia,
Romania,
Serbia,
Uruguay, or
Venezuela.
✻ OVERVIEW ✻
You will provide me with a very simple web scraper which I will use to scrape user names from Behance.net.
You may use whatever web scraper you prefer except for Selenium. You must *not* use Selenium.
I will run the web scraper on Lubuntu 20.04 LTS on my laptop computer.
The web scraper will append user names to a text file named, "behance_user_names.txt".
For example, behance_user_names.txt will contain data as follows...
https://www.behance.net/svitlnart
https://www.behance.net/demsey_dee
https://www.behance.net/katebel
and so on.
✻ DETAILS ✻
For example, please go to...
https://www.behance.net/search/projects?country=UA&field=illustration
Currently I obtain user names by right clicking on a user name and then choosing "Copy link address". Please see... https://i.imgur.com/8QIKnSX.png
Then I go to a URL such as this... https://www.behance.net/svitlnart so that I can review that illustrator's portfolio.
The phrase "plusieurs propriétaires is French for "multiple owners" in English. Please see... https://i.imgur.com/6mowuJ5.jpg. The web scraper will ignore URLs when a project has more than one owner (plusieurs propriétaires or multiple owners). In other words, I don't need URLs when a project has more than one owner (plusieurs propriétaires or multiple owners). But I need URLs when a project has one owner.
The web scraper must scrape all of the URLs on a page for projects that have one owner. What do I mention this? Isn't this obvious? I want to ensure you understand this requirement. Behance.net uses infinite scroll https://en.wiktionary.org/wiki/infinite_scroll. Therefore, the web scraper must continue to scroll down each page and scrape user names until the end of the page. I want all of the "infinite" results on the page I am scrapping. Obviously, the results aren't infinite; the results are obviously finite.
✻ SCREENCAST ✻
In addition to the web scraper, you will provide me with a screencast that is between one minute and ten minutes.
The screencast does *not* need audio. (You don't need to narrate the screencast). However, if you prefer, you may narrate the screencast. Are you confused? If you don't want to speak English then you don't need to speak English. See, I realize many programmers are uncomfortable speaking English.
The screencast will clearly explain to me how to use the web scraper.
The screencast will also show me how to install all necessary applications necessary to use the web scraper on Linux. (I use Lubuntu 20.04 LTS).
The screencast must *not* be created on Windows or Mac. The screencast must be created on Linux.
The screencast must not be larger than 50 megabytes yet the screencast must be easy for a person with normal vision to understand. In other words, you must not compress the screencast so much that it is difficult for a person with normal vision to understand.
✻ HOW TO APPLY ✻
If you are interested in working on this project with me then please copy and paste...
*** I am interested in this project to scrape data from Behance.net. Due to a glitch with Freelancer, I might not receive an email notification from Freelancer indicating that you have responded to me. Therefore, I might need to login to Freelancer to find out that you have responded to me.***
...into a message you send me. (Please do not send me the three asterisks before and after the text).
Please do *not* bother wasting your time including any other information. Instead, simply copy and paste the information above.
Related categories:
Business, Accounting, Human Resources & Legal
JavaScript
Python
Web Scraping
Golang