Summary
Blu Omega is a technology and cybersecurity solutions provider supporting federal agencies and enterprise clients. The company is seeking a Data Engineer to design, develop, and maintain cloud-based data pipelines, platforms, and stores for a federal data modernization program, while supporting analytics, data quality, testing, and platform improvement.
Responsibilities
- Design, develop, and maintain scalable ETL/ELT pipelines to ingest, transform, and deliver structured and unstructured data from multiple cloud-based sources
- Build and optimize data storage solutions, including data warehouses, data lakes, and lakehouses, to support large-scale analytics, reporting, and business intelligence
- Develop and maintain scalable data stores to facilitate analysis and downstream applications
- Write SQL and Python to retrieve, parse, transform, and process data, ensuring data quality and integrity
- Perform backend SQL and NoSQL testing, including validation, troubleshooting, and schema review
- Monitor and troubleshoot data pipelines to ensure reliable data ingestion and transformation
- Develop scripts and automation to standardize diverse datasets into analysis-ready formats
- Collaborate with analysts, developers, and stakeholders to translate requirements into technical data models and workflows
- Support the deployment, operation, and continuous improvement of cloud data platforms
- Leverage modern development practices and AI-assisted coding tools to enhance workflows
Skills
- 3+ years of experience in data warehouse development
- 3+ years of experience analyzing and integrating data within cloud-based environments
- 2+ years of experience developing and maintaining scalable data stores supporting big data analytics
- Hands-on experience using SQL and Python for data processing and transformation
- Experience designing and maintaining scalable ETL/ELT workflows using cloud technologies
- Familiarity with Cloudera, Hadoop, or related big data technologies
- Knowledge of cloud-based data architectures, storage, and pipelines
- Experience with AI-assisted coding tools supporting development
- Ability to perform backend SQL and NoSQL database testing and validation
- Strong analytical and troubleshooting skills
- Bachelor's degree in Computer Science, Data Science, Information Systems, Engineering, or related field
- Ability to obtain and sustain a Public Trust suitability determination
- Experience working within cross-functional engineering teams
- Familiarity with AWS data services such as AWS Glue, AWS Athena, and PostgreSQL
- Knowledge of distributed data processing technologies such as Spark, Databricks, Hive, AWS EMR, or Kafka
- Experience developing real-time data pipelines and streaming applications
- Working knowledge of NoSQL solutions like MongoDB or Cassandra
- Experience with data warehouse platforms such as AWS Redshift, MySQL, or Snowflake
- UNIX/Linux command-line and shell scripting proficiency
- Familiarity with Agile development practices
Qualifications
Must Haves
- 3+ years of experience in data warehouse development
- 3+ years of experience analyzing and integrating data within cloud-based environments
- 2+ years of experience developing and maintaining scalable data stores supporting big data analytics
- Hands-on experience using SQL and Python for data processing and transformation
- Experience designing and maintaining scalable ETL/ELT workflows using cloud technologies
- Familiarity with Cloudera, Hadoop, or related big data technologies
- Knowledge of cloud-based data architectures, storage, and pipelines
- Experience with AI-assisted coding tools supporting development
- Ability to perform backend SQL and NoSQL database testing and validation
- Strong analytical and troubleshooting skills
- Bachelor's degree in Computer Science, Data Science, Information Systems, Engineering, or related field
- Ability to obtain and sustain a Public Trust suitability determination
Nice to Haves
- Experience working within cross-functional engineering teams
- Familiarity with AWS data services such as AWS Glue, AWS Athena, and PostgreSQL
- Knowledge of distributed data processing technologies such as Spark, Databricks, Hive, AWS EMR, or Kafka
- Experience developing real-time data pipelines and streaming applications
- Working knowledge of NoSQL solutions like MongoDB or Cassandra
- Experience with data warehouse platforms such as AWS Redshift, MySQL, or Snowflake
- UNIX/Linux command-line and shell scripting proficiency
- Familiarity with Agile development practices
Benefits