Sequence Alignment in a High-Throughput Way
Budget: $250 – $750 USD
I am looking for a freelancer who can help with sequence alignment in a high-throughput way. The desired output of the sequence alignment is a list of aligned sequences.
Here is the description:
We have an excel spreadsheet with tens of thousands of test DNA sequences. We need to compare them to another DNA sequence (benchmark) and select test sequences that have a short sequence (15-18 bp) of high identity within them
A more detailed view:
- benchmark sequence. It is just one sequence for each project. Its length can wary from 300bp to 5000bp
- Tested sequences are all different. typically 30-50 bp, up to 100pb
- need to identify test sequences that have at least one stretch of 15-18 bp with at least 50% identity to the benchmark
Ideally, we would like to have a system that we can use ourselves in our lab but we are open for a service as well. Please let me know if that is something you can help us with.
The project has a test file attached. Please contact me only if you can align at least 1 out of 3 test sequences and the benchmark sequence. We can discuss the real data after I make sure that you can do it.
Here is the description:
We have an excel spreadsheet with tens of thousands of test DNA sequences. We need to compare them to another DNA sequence (benchmark) and select test sequences that have a short sequence (15-18 bp) of high identity within them
A more detailed view:
- benchmark sequence. It is just one sequence for each project. Its length can wary from 300bp to 5000bp
- Tested sequences are all different. typically 30-50 bp, up to 100pb
- need to identify test sequences that have at least one stretch of 15-18 bp with at least 50% identity to the benchmark
Ideally, we would like to have a system that we can use ourselves in our lab but we are open for a service as well. Please let me know if that is something you can help us with.
The project has a test file attached. Please contact me only if you can align at least 1 out of 3 test sequences and the benchmark sequence. We can discuss the real data after I make sure that you can do it.