Genomic Data Curation Specialists Needed

Job ID: 40034677

Budget: $15 – $25 USD

We have a lot of publicly available genomics data that need to be downloaded in the cloud to run our workflows and curate the data the guidelines. The scope covers three tightly-linked tasks: data cleaning, data annotation, and data integration. In practice, this means screening FASTQ/BAM files for quality issues, enriching each sample with consistent metadata, then merging the curated sets so they slot straight into my downstream analytics pipeline.

I already run an in-house pipeline, so you’ll be working with the specific software I provide (all open source). If you are comfortable adapting to established nextflow workflows in genomics, you should have no trouble.

Deliverables
• A cleaned, quality-controlled sequencing dataset
• Structured annotation files that match my schema exactly
• An integrated master dataset packaged for immediate analysis

I’ll supply sample data, the curation guidelines, and access to the tools as soon as we agree on milestones. Looking forward to collaborating with someone who has an eye for detail and a solid grasp of sequencing-based data management.