Summary
Ariadne is seeking a Data Engineer to support its work at the National Institutes of Health. The role focuses on operating and troubleshooting data pipelines, working with relational databases and scripting languages, and supporting bioinformatics databases and workflows in a complex technical environment.
Skills
- B.S. in a STEM field (Engineering, Computer Science, Mathematics, Physics)
- Alternatively, equivalent industry experience in bioinformatics or a related field
- Experience running operations in a large and complex environment, preferably in data operations
- Relational databases, SQL
- Scripting in Bash, Python, or other shell scripting languages
- Experience with LINUX/UNIX
- Ability to troubleshoot an operational pipeline to identify highest priority problems and identify solutions
- Excellent interpersonal skills and the ability to work as part of a team
- Knowledge of existing workflow languages and frameworks
- Work experience with production-level bioinformatics databases and pipelines
- Familiarity with technical environments, complex databases, and process flows
- Experience with the NCBI Sequence Read Archive (SRA) or GenBank databases and tools like BLAST or other DNA sequence analysis software
- Cloud technologies like Kubernetes, Athena, and BigQuery
- Experience with XML schemas
- Familiarity with Jira and Confluence
- Experience with Agile processes, especially scrum
- Background in virology is a plus
- Strong presentation skills
Qualifications
Must Haves
- B.S. in a STEM field (Engineering, Computer Science, Mathematics, Physics)
- Alternatively, equivalent industry experience in bioinformatics or a related field
- Experience running operations in a large and complex environment, preferably in data operations
- Relational databases, SQL
- Scripting in Bash, Python, or other shell scripting languages
- Experience with LINUX/UNIX
- Ability to troubleshoot an operational pipeline to identify highest priority problems and identify solutions
- Excellent interpersonal skills and the ability to work as part of a team
Nice to Haves
- Knowledge of existing workflow languages and frameworks
- Work experience with production-level bioinformatics databases and pipelines
- Familiarity with technical environments, complex databases, and process flows
- Experience with the NCBI Sequence Read Archive (SRA) or GenBank databases and tools like BLAST or other DNA sequence analysis software
- Cloud technologies like Kubernetes, Athena, and BigQuery
- Experience with XML schemas
- Familiarity with Jira and Confluence
- Experience with Agile processes, especially scrum
- Background in virology is a plus
- Strong presentation skills
Benefits
- Full-time opportunity at the NIH in Bethesda, MD and/or remote work