Python Developer - Data Mining
Budget: £20 – £250 GBP
Job Title
Python Developer - Data Mining
About the Company
You’ll be working with ‘The Classic Valuer’ (TCV). TCV is focused on classic cars. More specifically, helping people know the price and price trend of every single classic car in the world.
We do that by collating every time a car has sold at auction into one place and then analysing the results.
We’re a start-up. Small and nimble. You’ll be working directly with the two founders, Charlie and Giles - both of whom are UK based.
Job Purpose
Put simply, you’ll be responsible for the most important part of the website - scraping data from auction sites around the world and passing them through our data pipelines to cleanse and sanitise the data.
To date, Charlie has been responsible for the above. He’ll be your point of contact to provide guidance on our existing approach and answer questions as needed.
Job Duties and Responsibilities
We are looking for someone to be responsible for developing and maintaining our data ingestion engine to bring data from auction houses into our database.
Written in Python, this makes heavy use of the Scrapy web framework, in addition to custom modules. As a fully remote team, we use version control and other DevOps tools to collaborate.
You will be responsible for:
Building web scrapers for additional data sources, using TCV libraries and methodology.
Maintaining existing web scrapers and data pipelines
Performing data quality checks on newly scraped data
Using Git to collaborate with other developers, integrating any changes into the main code base following a peer review process
Skills / Experience
In addition to the technical skills below, the ideal candidate will be an excellent problem solver and have a proactive attitude towards finding and fixing bugs.
Requirements:
1+ years of Python experience, working on large code bases
Good knowledge of Git, and ability to apply best practice principles
Experience using Scrapy and Selenium for data mining in Python
Nice to have:
Experience of working within a team of developers
Experience using Docker
Experience using Cloud technologies (AWS / GCP)
Experience using Unix based operating systems (build tools are built for Bash)
Ways of Working
We operate on a results basis, we won’t be micromanaging your time. Rather we will be setting targets (e.g. how many auction sites to scrape, data quality and accuracy etc) and supporting you as needed to deliver on those targets.
We will work remotely, so you can be based anywhere in the world.
Python Developer - Data Mining
About the Company
You’ll be working with ‘The Classic Valuer’ (TCV). TCV is focused on classic cars. More specifically, helping people know the price and price trend of every single classic car in the world.
We do that by collating every time a car has sold at auction into one place and then analysing the results.
We’re a start-up. Small and nimble. You’ll be working directly with the two founders, Charlie and Giles - both of whom are UK based.
Job Purpose
Put simply, you’ll be responsible for the most important part of the website - scraping data from auction sites around the world and passing them through our data pipelines to cleanse and sanitise the data.
To date, Charlie has been responsible for the above. He’ll be your point of contact to provide guidance on our existing approach and answer questions as needed.
Job Duties and Responsibilities
We are looking for someone to be responsible for developing and maintaining our data ingestion engine to bring data from auction houses into our database.
Written in Python, this makes heavy use of the Scrapy web framework, in addition to custom modules. As a fully remote team, we use version control and other DevOps tools to collaborate.
You will be responsible for:
Building web scrapers for additional data sources, using TCV libraries and methodology.
Maintaining existing web scrapers and data pipelines
Performing data quality checks on newly scraped data
Using Git to collaborate with other developers, integrating any changes into the main code base following a peer review process
Skills / Experience
In addition to the technical skills below, the ideal candidate will be an excellent problem solver and have a proactive attitude towards finding and fixing bugs.
Requirements:
1+ years of Python experience, working on large code bases
Good knowledge of Git, and ability to apply best practice principles
Experience using Scrapy and Selenium for data mining in Python
Nice to have:
Experience of working within a team of developers
Experience using Docker
Experience using Cloud technologies (AWS / GCP)
Experience using Unix based operating systems (build tools are built for Bash)
Ways of Working
We operate on a results basis, we won’t be micromanaging your time. Rather we will be setting targets (e.g. how many auction sites to scrape, data quality and accuracy etc) and supporting you as needed to deliver on those targets.
We will work remotely, so you can be based anywhere in the world.