Ariadne logo
Ariadne
Posted 22 days agoVerified live 14h ago

NCBI Data Engineer

Brief overview

Remote
UndergradOr in progress
3+ yrsMinimum
Data OperationsRelational DatabasesSQLBashPythonLinux/UNIXOperational Pipeline TroubleshootingBioinformatics Databases and PipelinesNCBI Sequence Read Archive (SRA)GenBankBLASTKubernetes

About the company

Ariadne is a high technology company that provides people counting and customer flow analytics with its patented cutting edge AI technology.

Job description

Summary

Ariadne is seeking a Data Engineer to support its work at the National Institutes of Health. The role focuses on operating and troubleshooting data pipelines, working with relational databases and scripting languages, and supporting bioinformatics databases and workflows in a complex technical environment.

Skills

  • B.S. in a STEM field (Engineering, Computer Science, Mathematics, Physics)
  • Alternatively, equivalent industry experience in bioinformatics or a related field
  • Experience running operations in a large and complex environment, preferably in data operations
  • Relational databases, SQL
  • Scripting in Bash, Python, or other shell scripting languages
  • Experience with LINUX/UNIX
  • Ability to troubleshoot an operational pipeline to identify highest priority problems and identify solutions
  • Excellent interpersonal skills and the ability to work as part of a team
  • Knowledge of existing workflow languages and frameworks
  • Work experience with production-level bioinformatics databases and pipelines
  • Familiarity with technical environments, complex databases, and process flows
  • Experience with the NCBI Sequence Read Archive (SRA) or GenBank databases and tools like BLAST or other DNA sequence analysis software
  • Cloud technologies like Kubernetes, Athena, and BigQuery
  • Experience with XML schemas
  • Familiarity with Jira and Confluence
  • Experience with Agile processes, especially scrum
  • Background in virology is a plus
  • Strong presentation skills

Qualifications

Must Haves

  • B.S. in a STEM field (Engineering, Computer Science, Mathematics, Physics)
  • Alternatively, equivalent industry experience in bioinformatics or a related field
  • Experience running operations in a large and complex environment, preferably in data operations
  • Relational databases, SQL
  • Scripting in Bash, Python, or other shell scripting languages
  • Experience with LINUX/UNIX
  • Ability to troubleshoot an operational pipeline to identify highest priority problems and identify solutions
  • Excellent interpersonal skills and the ability to work as part of a team

Nice to Haves

  • Knowledge of existing workflow languages and frameworks
  • Work experience with production-level bioinformatics databases and pipelines
  • Familiarity with technical environments, complex databases, and process flows
  • Experience with the NCBI Sequence Read Archive (SRA) or GenBank databases and tools like BLAST or other DNA sequence analysis software
  • Cloud technologies like Kubernetes, Athena, and BigQuery
  • Experience with XML schemas
  • Familiarity with Jira and Confluence
  • Experience with Agile processes, especially scrum
  • Background in virology is a plus
  • Strong presentation skills

Benefits

  • Full-time opportunity at the NIH in Bethesda, MD and/or remote work

More jobs like this