G
Georgia IT, Inc.
Posted 31 days agoVerified live 15h ago

Data Engineer

Brief overview

Remote
4+ yrsMinimum
Data Processing SystemsSparkAirflowSQLETL FrameworksFlumeOozieSpark ScalaDistributed StorageS3HiveJavaPythonGitHub

About the company

G
Georgia IT, Inc.georgiait.com

Georgia IT, Inc. provides IT Consulting for a wide range of IT services and custom build turn-key enterprise solutions.

Job description

Summary

Georgia IT, Inc. is seeking a Data Engineer for a remote contract position. The role involves developing and automating large scale data processing systems to enhance business growth and improve product experience.

Responsibilities

  • Develop and automate large scale, high-performance data processing systems (batch and/or streaming) to drive Airbnb business growth and improve the product experience
  • Build scalable Spark data pipelines leveraging Airflow scheduler/executor framework

Skills

  • 4+ years of relevant industry experience
  • Demonstrated ability to analyze large data sets to identify gaps and inconsistencies, provide data insights, and advance effective product solutions
  • Working knowledge of relational databases and query authoring (SQL)
  • Good communication skills, both written and verbal
  • Strong experience using ETL framework (ex: Airflow, Flume, Oozie etc.) to build and deploy production-quality ETL pipelines
  • Experience building batch data pipelines in Spark Scala
  • Strong understanding of distributed storage and compute (S3, Hive, Spark)
  • General software engineering skills (Java or Python, Github)
  • US Citizen/ GC /EAD (H4/L2/TN) preferred- No 3rdPARTIES RESUME C2C ACCEPTED

Qualifications

Must Haves

  • 4+ years of relevant industry experience
  • Demonstrated ability to analyze large data sets to identify gaps and inconsistencies, provide data insights, and advance effective product solutions
  • Working knowledge of relational databases and query authoring (SQL)
  • Good communication skills, both written and verbal
  • Strong experience using ETL framework (ex: Airflow, Flume, Oozie etc.) to build and deploy production-quality ETL pipelines
  • Experience building batch data pipelines in Spark Scala
  • Strong understanding of distributed storage and compute (S3, Hive, Spark)
  • General software engineering skills (Java or Python, Github)

Nice to Haves

  • US Citizen/ GC /EAD (H4/L2/TN) preferred- No 3rdPARTIES RESUME C2C ACCEPTED

Benefits

  • Work_model: [remote]
  • Job Name : Data Engineer - REMOTE

More jobs like this