Website Data Scraping: Public Records

Job ID: 37805036

Budget: $2 – $8 USD

There is data on a website that I need to have scraped. I have prepared detailed instructions. You will log in to the site, copy/paste data into a spreadsheet, and save certain files into a Google Drive folder.

Below is the text of the directions but once you are hired I will share the Google Sheet with you which has the instructions more neatly formatted, links to the Google Drive folder, and login information for the site.

---

1 Go to (link removed)
2 Log in as user: [TO BE PROVIDED]
3 Go to "View My Requests"
4 Start from page 1 and add each request to the "Data" tab on this spreadsheet.
5 First, copy and paste the record ID (this is the text right under "Open Records Request", like "C000374-011124". It should paste as a link.
6 Check on the right side to see "In Progress" or "Complete" and copy that into "Progress".
7 Look for Status and copy that text into the "Status" field (e.g. No Records Exist, Sent to Attorney General, or whatever it says).
8 Copy the text of the request (starting under the the Record ID and ending right above Status) and paste into "Request Text".
9 If there is NOT a button under the request that says "View Files" then put 0 for "# of Files" and leave "Google Drive Link" blank.
10 Go to the next record.

If the record DOES have a "View File(s)" Button:

11 Click on "View File(s). On the next page, count how many files there are and add that number to "# of Files".
12 Go to this Google Drive location and create a new folder with just the record ID number. Save all the files in there.
NOTE: If you "Download All" to save time, you must unzip the files and upload them all to Google Drive. Do not upload a .ZIP file.
13 Go to the next record.
Related categories: Data Entry Web Scraping Data Mining Google Docs