Extract the clean document from Resume Data with python
Budget: $10 – $30 USD
Extract the clean document from Resume Data with python
- Purpose
Extracting the clean document set from the original resume data set for machine learning.
- Target Resume
resume_samples.zip
https://github.com/florex/resume_corpus
- Expected output:
+ Deliver the output data as a CSV file and program code in python with appropriate comments in the code
+ Version: python 3.6* or higher
+ Clean document with encoding in UTF-8
+ Delete stop words, HTML-related tags, file paths
+ 'C:\\\Workspace\\\java\\\scrape_indeed\\\.*.html' need to be deleted.
+ Extract the clean document as the enumerated words groups in one column for each document in the CSV file
(the CSV has two columns that are "Original Resume" and "Cleaned Resume")
- Deadline
1-day
- Notice
there is the possibility to ask you about the continuous work in python programming.
This is a simple task. However, please show us your ability in coding to confirm your programming skill.
- Purpose
Extracting the clean document set from the original resume data set for machine learning.
- Target Resume
resume_samples.zip
https://github.com/florex/resume_corpus
- Expected output:
+ Deliver the output data as a CSV file and program code in python with appropriate comments in the code
+ Version: python 3.6* or higher
+ Clean document with encoding in UTF-8
+ Delete stop words, HTML-related tags, file paths
+ 'C:\\\Workspace\\\java\\\scrape_indeed\\\.*.html' need to be deleted.
+ Extract the clean document as the enumerated words groups in one column for each document in the CSV file
(the CSV has two columns that are "Original Resume" and "Cleaned Resume")
- Deadline
1-day
- Notice
there is the possibility to ask you about the continuous work in python programming.
This is a simple task. However, please show us your ability in coding to confirm your programming skill.