Summary
Mizuho is a leading financial services company, and they are seeking a Data Engineer to develop data solutions for the enterprise. The role involves building and maintaining data pipelines and collaborating with teams to ensure accurate and scalable data solutions.
Responsibilities
- Assist in building and maintaining ingestion pipelines that land raw data into the Bronze layer
- Support Silver layer transformations under guidance: cleansing, deduplication, and schema enforcement
- Write SQL and PySpark for defined transformation tasks
- Run and monitor scheduled jobs; help investigate and resolve pipeline failures
- Document pipeline logic, transformations, and fixes
- Participate in code reviews as a reviewer-in-training and incorporate feedback on your own work
- Learn team standards for version control, testing, and deployment
Skills
- 0-2 years of experience in data engineering, analytics, or a related technical role (internships and academic projects count)
- Foundational SQL skills (joins, aggregations, filtering)
- Python proficiency with basic OOPS knowledge
- Understanding of core data concepts (tables, schemas, relational data)
- Willingness to learn Databricks, Spark, and cloud technologies
- Familiarity with Git or a demonstrated ability to learn version control quickly
- Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
- Proven experience in data engineering, software development, or related roles
- Proficiency in programming languages commonly used in data engineering (e.g., Python, Scala, etc.)
- Strong knowledge of database systems, data modeling techniques, and SQL proficiency
- Proficiency with ETL tools commonly used in data engineering (e.g., SSIS, Databricks, Azure Data Factory)
- Experience with big data technologies and frameworks (e.g., Spark, Kafka, etc.)
- Familiarity with cloud platforms and services (e.g., Azure)
- Excellent problem-solving skills and attention to detail
- Effective communication and collaboration skills in a team-oriented environment
- Ability to adapt to evolving technologies and business requirements
- Exposure to Databricks, Apache Spark, or PySpark (coursework or hands-on)
- Awareness of the medallion architecture and Delta Lake basics
- Experience with any cloud platform (Azure, AWS, or Google Cloud Platform)
- Relevant coursework, bootcamp, or a Databricks certification (e.g., Data Engineer Associate)
- Any experience with data visualization or BI tools
Qualifications
Must Haves
- 0-2 years of experience in data engineering, analytics, or a related technical role (internships and academic projects count)
- Foundational SQL skills (joins, aggregations, filtering)
- Python proficiency with basic OOPS knowledge
- Understanding of core data concepts (tables, schemas, relational data)
- Willingness to learn Databricks, Spark, and cloud technologies
- Familiarity with Git or a demonstrated ability to learn version control quickly
- Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
- Proven experience in data engineering, software development, or related roles
- Proficiency in programming languages commonly used in data engineering (e.g., Python, Scala, etc.)
- Strong knowledge of database systems, data modeling techniques, and SQL proficiency
- Proficiency with ETL tools commonly used in data engineering (e.g., SSIS, Databricks, Azure Data Factory)
- Experience with big data technologies and frameworks (e.g., Spark, Kafka, etc.)
- Familiarity with cloud platforms and services (e.g., Azure)
- Excellent problem-solving skills and attention to detail
- Effective communication and collaboration skills in a team-oriented environment
- Ability to adapt to evolving technologies and business requirements
Nice to Haves
- Exposure to Databricks, Apache Spark, or PySpark (coursework or hands-on)
- Awareness of the medallion architecture and Delta Lake basics
- Experience with any cloud platform (Azure, AWS, or Google Cloud Platform)
- Relevant coursework, bootcamp, or a Databricks certification (e.g., Data Engineer Associate)
- Any experience with data visualization or BI tools
Benefits
- A generous employee benefits package, including Medical, Dental and 401K plans
- Successful candidates are also eligible to receive a discretionary bonus