Design, develop, and productionize scalable bioinformatics workflows supporting pangenome graph construction, haplotype expansion, population-scale imputation, and genomic data quality control. This role focuses on transforming research-grade analyses into robust, reusable, and cloud-enabled pipelines that support large-scale multi-crop genomics programs. Develop and maintain production-grade bioinformatics workflows for pangenome, haplotype, imputation, and QC processes. Convert research scripts and manual analyses into automated, version-controlled, and reproducible pipelines. Work with genomic data formats including FASTA, GFF/GTF, VCF, BAM/CRAM, haplotype outputs, and associated metadata. Implement workflows using cloud-native AWS services, leveraging S3 storage and scalable batch execution. Build validation, logging, provenance tracking, and error-handling capabilities into workflows. Collaborate with scientists and domain experts to ensure biological accuracy and usability of outputs.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
No Education Listed