I will build a reproducible bioinformatics pipeline in nextflow
Level 1
Has met certain performance criteria and shows strong potential in the marketplace.
About this Gig
I build bioinformatics pipelines in Nextflow, either by adapting an analysis you're already running manually, or by designing one from scratch if you're starting from just a set of tools and a goal.
Nextflow is the right choice here for most labs: it handles parallelization and resuming failed runs without extra configuration, and it runs the same way locally, on a cluster, or in the cloud, so the pipeline doesn't need to be rebuilt if your compute situation changes later.
If you have an existing workflow, I'll restructure it into defined processes with clear inputs and outputs at each stage, and a Conda environment or container per process so tool versions are fixed and the results are reproducible. If you're starting from scratch, we'll work out the steps together based on your data type and analysis goals before I build anything.
I will write the documentation to use it assuming the person running the pipeline isn't a bioinformatician and may not know Nextflow at all. That means plain-language setup instructions, what each parameter actually controls (not just its name), what a successful run looks like, and what to do with the most common failure messages.
Expertise:
Automations
Technology:
Other
FAQ
Do I need to already have a working analysis, or can you design one from scratch?
Either. I can formalize an existing workflow, or design the analysis with you if you're starting from just a research question and a dataset.
