Dice logo
Dice
Posted 57 days agoVerified live 12h ago

Data Engineer

Brief overview

Remote
UndergradOr in progress
$88k–$105k/yrStated range
SQLPythonObject-Oriented ProgrammingDatabricksApache SparkPySparkGitAzure CloudData ModelingETL ToolsKafka

About the company

Dice is the go-to career marketplace for tech professionals.

Job description

Summary

Mizuho is a leading financial services company, and they are seeking a Data Engineer to develop data solutions for the enterprise. The role involves building and maintaining data pipelines and collaborating with teams to ensure accurate and scalable data solutions.

Responsibilities

  • Assist in building and maintaining ingestion pipelines that land raw data into the Bronze layer
  • Support Silver layer transformations under guidance: cleansing, deduplication, and schema enforcement
  • Write SQL and PySpark for defined transformation tasks
  • Run and monitor scheduled jobs; help investigate and resolve pipeline failures
  • Document pipeline logic, transformations, and fixes
  • Participate in code reviews as a reviewer-in-training and incorporate feedback on your own work
  • Learn team standards for version control, testing, and deployment

Skills

  • 0-2 years of experience in data engineering, analytics, or a related technical role (internships and academic projects count)
  • Foundational SQL skills (joins, aggregations, filtering)
  • Python proficiency with basic OOPS knowledge
  • Understanding of core data concepts (tables, schemas, relational data)
  • Willingness to learn Databricks, Spark, and cloud technologies
  • Familiarity with Git or a demonstrated ability to learn version control quickly
  • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
  • Proven experience in data engineering, software development, or related roles
  • Proficiency in programming languages commonly used in data engineering (e.g., Python, Scala, etc.)
  • Strong knowledge of database systems, data modeling techniques, and SQL proficiency
  • Proficiency with ETL tools commonly used in data engineering (e.g., SSIS, Databricks, Azure Data Factory)
  • Experience with big data technologies and frameworks (e.g., Spark, Kafka, etc.)
  • Familiarity with cloud platforms and services (e.g., Azure)
  • Excellent problem-solving skills and attention to detail
  • Effective communication and collaboration skills in a team-oriented environment
  • Ability to adapt to evolving technologies and business requirements
  • Exposure to Databricks, Apache Spark, or PySpark (coursework or hands-on)
  • Awareness of the medallion architecture and Delta Lake basics
  • Experience with any cloud platform (Azure, AWS, or Google Cloud Platform)
  • Relevant coursework, bootcamp, or a Databricks certification (e.g., Data Engineer Associate)
  • Any experience with data visualization or BI tools

Qualifications

Must Haves

  • 0-2 years of experience in data engineering, analytics, or a related technical role (internships and academic projects count)
  • Foundational SQL skills (joins, aggregations, filtering)
  • Python proficiency with basic OOPS knowledge
  • Understanding of core data concepts (tables, schemas, relational data)
  • Willingness to learn Databricks, Spark, and cloud technologies
  • Familiarity with Git or a demonstrated ability to learn version control quickly
  • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
  • Proven experience in data engineering, software development, or related roles
  • Proficiency in programming languages commonly used in data engineering (e.g., Python, Scala, etc.)
  • Strong knowledge of database systems, data modeling techniques, and SQL proficiency
  • Proficiency with ETL tools commonly used in data engineering (e.g., SSIS, Databricks, Azure Data Factory)
  • Experience with big data technologies and frameworks (e.g., Spark, Kafka, etc.)
  • Familiarity with cloud platforms and services (e.g., Azure)
  • Excellent problem-solving skills and attention to detail
  • Effective communication and collaboration skills in a team-oriented environment
  • Ability to adapt to evolving technologies and business requirements

Nice to Haves

  • Exposure to Databricks, Apache Spark, or PySpark (coursework or hands-on)
  • Awareness of the medallion architecture and Delta Lake basics
  • Experience with any cloud platform (Azure, AWS, or Google Cloud Platform)
  • Relevant coursework, bootcamp, or a Databricks certification (e.g., Data Engineer Associate)
  • Any experience with data visualization or BI tools

Benefits

  • A generous employee benefits package, including Medical, Dental and 401K plans
  • Successful candidates are also eligible to receive a discretionary bonus

More jobs like this