Summary
Protagona is a company that specializes in cloud resources deployment and configuration. They are seeking a Data Engineer to join their team, responsible for creating and managing data pipelines and ensuring the effective use of cloud technologies to meet client needs.
Responsibilities
- Work with the team to evaluate business needs and priorities, liaise with key business partners and address team needs related to data systems and management
- Translate business requirements into technical specifications; establish and define details, definitions, and requirements of applications, components and enhancements
- Participate in project planning; identifying milestones, deliverables and resource requirements; tracks activities and task execution
- Generate design, development, test plans, detailed functional specifications documents, user interface design, and process flow charts for execution of programming
- Develop data pipelines / APIs using Python, SQL, potentially Spark and AWS, Azure or GCP Methods
- Use an analytical, data-driven approach to drive a deep understanding of fast changing business
- Build large-scale batch and real-time data pipelines with data processing frameworks in AWS, Azure or GCP cloud platform
- Moving data from on-prem to cloud and cloud data conversions
Skills
- Experience in data engineering with an emphasis on data analytics and reporting
- Exposure to the AWS Cloud Platform
- Experience in SQL, data transformations, and troubleshooting across at least one database Platform (Redshift, Amazon RDS, Cassandra, Snowflake, PostgreSQL, Databricks, etc.)
- Experience in the design and build of data extraction, transformation, and loading processes by writing custom data pipelines
- Experience in a scripting language such as Python
- Experience designing and building solutions utilizing various Cloud services such as EC2, S3, EMR, Kinesis, RDS, Redshift/Spectrum, Lambda, Glue, Athena, API gateway, etc
Qualifications
Must Haves
- Experience in data engineering with an emphasis on data analytics and reporting
- Exposure to the AWS Cloud Platform
- Experience in SQL, data transformations, and troubleshooting across at least one database Platform (Redshift, Amazon RDS, Cassandra, Snowflake, PostgreSQL, Databricks, etc.)
- Experience in the design and build of data extraction, transformation, and loading processes by writing custom data pipelines
- Experience in a scripting language such as Python
- Experience designing and building solutions utilizing various Cloud services such as EC2, S3, EMR, Kinesis, RDS, Redshift/Spectrum, Lambda, Glue, Athena, API gateway, etc