Summary
Georgia IT, Inc. is seeking a Data Engineer who will work with the business to understand data requirements and become a data platform expert in designing and building data solutions focused on Cloud-based Big Data ecosystems. The role involves collaborating with data science teams and taking ownership of projects to deliver high-quality data-driven advanced analytics applications.
Responsibilities
- Collaborate and work with global data management stakeholders to identify requirements for complex business problems that may be loosely defined
- Work with the business, applications owners, solutions architects, and with technical architects to understand proposed solution architectures
- Build, deploy and monitor Batch and near real time data pipelines to load structured and unstructured data into data lake platforms
- Identify, evaluate and implement leading edge data management frameworks required to address complex large-scale data challenges
- Work within multi-functional agile teams with end-to-end responsibility for product development and delivery
- Provide architectural support by building proof of concepts & prototypes
Skills
- Experienced in programming languages such as python, SQL and spark
- Experience or exposure Jupyter Notebook, etc
- Experience building Data engineering pipelines in corporate data lake environment handling large structured and unstructured datasets
- Good understanding of linux os, security and scripting
- Energetic, able to build and sustain long-term relationships across a multitude of stakeholders in a fast paced, multi-national work environment
- Strong time management and organizational skills
- Possess strong verbal and written communication skills and ability to present, persuade and influence peers
- Bachelor's degree in Information systems or related field with GPA of 3.0+ required
- Excellent data analysis skills
- Experience in performing analysis and design for data management and data driven projects
- Familiarity with data science and analytic tool sets
- Exposure or experience with Cloud Platforms, Azure, Databricks, SQLDW and COSMOSDB
- Experience in designing and leading the conceptual, logical and physical design for distributed databases
- Experience with operating system command languages such as bash or ksh
- Experience with development tools such as git and integrated development environments
- Understanding of the SAFe Agile development methodology
Qualifications
Must Haves
- Experienced in programming languages such as python, SQL and spark
- Experience or exposure Jupyter Notebook, etc
- Experience building Data engineering pipelines in corporate data lake environment handling large structured and unstructured datasets
- Good understanding of linux os, security and scripting
- Energetic, able to build and sustain long-term relationships across a multitude of stakeholders in a fast paced, multi-national work environment
- Strong time management and organizational skills
- Possess strong verbal and written communication skills and ability to present, persuade and influence peers
- Bachelor's degree in Information systems or related field with GPA of 3.0+ required
- Excellent data analysis skills
- Experience in performing analysis and design for data management and data driven projects
- Familiarity with data science and analytic tool sets
Nice to Haves
- Exposure or experience with Cloud Platforms, Azure, Databricks, SQLDW and COSMOSDB
- Experience in designing and leading the conceptual, logical and physical design for distributed databases
- Experience with operating system command languages such as bash or ksh
- Experience with development tools such as git and integrated development environments
- Understanding of the SAFe Agile development methodology