OVA.Work logo
OVA.Work
Posted 59 days agoVerified live 1d ago

Java Data Engineer

Brief overview

Remote
UndergradOr in progress
3+ yrsMinimum
Java 8/11/17Spring BootSQLRelational databasesNoSQL databasesREST APIsLinux/UnixGitAgile/ScrumApache SparkHadoopApache FlinkApache KafkaApache AirflowData lakesData warehousesDocker

About the company

OVA.Work logo
OVA.Workova.work

OVA is the most advanced Automated, Intelligent, intuitive On-boarding platform for Staffing Firms of all sizes.

Job description

Summary

OVA.Work is seeking a highly motivated Java Data Engineer to design, develop, and maintain scalable data pipelines and distributed data processing systems. The role involves collaborating with data scientists and software engineers to build robust data solutions that enable analytics and business intelligence.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Java
  • Build and optimize ETL/ELT workflows for processing large datasets
  • Develop high-performance data ingestion, transformation, and integration solutions
  • Design and maintain distributed data processing applications using Spark or Hadoop
  • Develop RESTful APIs for data services and integrations
  • Optimize SQL queries and database performance
  • Implement data quality, validation, and monitoring processes
  • Collaborate with cross-functional teams to understand business data requirements
  • Troubleshoot production data issues and optimize pipeline performance
  • Ensure data security, governance, and compliance standards are followed
  • Participate in Agile development, code reviews, and architecture discussions
  • Maintain technical documentation and best practices

Skills

  • Bachelor's degree in Computer Science, Information Technology, Data Engineering, or a related field
  • 3+ years of experience in Java development
  • Experience with data engineering or ETL development
  • Strong proficiency in Java (Java 8/11/17)
  • Experience with Spring Boot
  • Strong SQL and database design skills
  • Experience working with relational and NoSQL databases
  • Experience building REST APIs
  • Familiarity with Linux/Unix environments
  • Experience with Git and version control
  • Knowledge of Agile/Scrum methodologies
  • Experience with Apache Spark, Hadoop, or Apache Flink
  • Knowledge of Kafka or other messaging platforms
  • Experience with cloud platforms (AWS, Azure, or GCP)
  • Experience with data lakes and data warehouses
  • Familiarity with containerization using Docker and Kubernetes
  • Experience with Airflow or workflow orchestration tools
  • Knowledge of data modeling and dimensional modeling
  • Experience with CI/CD pipelines
  • Experience building enterprise-scale data platforms
  • Experience processing structured and unstructured data
  • Knowledge of streaming data architectures
  • Experience optimizing large-scale data pipelines
  • Familiarity with data governance and security practices
  • Experience working in Agile development teams
  • Apache Beam
  • Delta Lake
  • Databricks
  • Iceberg
  • Hudi
  • Terraform
  • Dbt
  • OAuth/JWT Authentication
  • Microservices Architecture
  • Event-Driven Architecture
  • Machine Learning data pipelines
  • Real-time analytics platforms

Qualifications

Must Haves

  • Bachelor's degree in Computer Science, Information Technology, Data Engineering, or a related field
  • 3+ years of experience in Java development
  • Experience with data engineering or ETL development
  • Strong proficiency in Java (Java 8/11/17)
  • Experience with Spring Boot
  • Strong SQL and database design skills
  • Experience working with relational and NoSQL databases
  • Experience building REST APIs
  • Familiarity with Linux/Unix environments
  • Experience with Git and version control
  • Knowledge of Agile/Scrum methodologies

Nice to Haves

  • Experience with Apache Spark, Hadoop, or Apache Flink
  • Knowledge of Kafka or other messaging platforms
  • Experience with cloud platforms (AWS, Azure, or GCP)
  • Experience with data lakes and data warehouses
  • Familiarity with containerization using Docker and Kubernetes
  • Experience with Airflow or workflow orchestration tools
  • Knowledge of data modeling and dimensional modeling
  • Experience with CI/CD pipelines
  • Experience building enterprise-scale data platforms
  • Experience processing structured and unstructured data
  • Knowledge of streaming data architectures
  • Experience optimizing large-scale data pipelines
  • Familiarity with data governance and security practices
  • Experience working in Agile development teams
  • Apache Beam
  • Delta Lake
  • Databricks
  • Iceberg
  • Hudi
  • Terraform
  • dbt
  • OAuth/JWT Authentication
  • Microservices Architecture
  • Event-Driven Architecture
  • Machine Learning data pipelines
  • Real-time analytics platforms

More jobs like this