Diverse Lynx logo
Diverse Lynx
Posted 3 days agoVerified live 15h ago

Data Engineer

Brief overview

Remote
UndergradOr in progress
$110k–$130k/yrStated range
3+ yrsMinimum
1 H-1B approvalsDept. of Labor
Google Cloud PlatformPySparkPythonApache AirflowApache HiveApache KafkaData Pipeline DevelopmentETL/ELT Pipeline Development

About the company

Diverse Lynx logo
Diverse Lynxdiverselynx.com

Diverse Lynx is a WBENC- and NMSDC-certified partner, helping organizations turn diversity goals into measurable impact through staffing and contingent workforce solutions.

Visa sponsorship history

1 year sponsoring, last filed FY2025

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
1H-1B approved
100%approval rate
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
20251

Job description

Summary

Confidential company is seeking a Data Engineer to support secure, reliable, and scalable data platforms and pipelines. The role focuses on developing ETL/ELT pipelines, automating data ingestion, optimizing PySpark processing, orchestrating workflows, and supporting analytics, reporting, AI/ML, and business operations.

Responsibilities

  • Data Pipeline Development, Enhancement & Maintenance
  • Design, build, and maintain scalable ETL/ELT pipelines
  • Automate data ingestion from multiple data sources
  • Ensure reliable, high-performance, and scalable data movement
  • Develop and Eligible to workimize PySpark-based data processing frameworks
  • Build and manage Airflow workflows and orchestration pipelines
  • Work with Hive and Kafka-based batch and streaming data solutions
  • Implement data quality, monitoring, governance, and operational controls
  • Support production incidents, performance tuning, and continuous improvement initiatives
  • The Data Engineer will be responsible for designing, building, governing, and Eligible to workimizing data platforms, pipelines, and integrations to ensure secure, reliable, and scalable data availability for analytics, reporting, AI/ML, and business operations

Skills

  • Google Cloud Platform
  • PySpark
  • Python
  • Airflow
  • Hive
  • Kafka

Qualifications

Must Haves

  • Google Cloud Platform
  • PySpark
  • Python
  • Airflow
  • Hive
  • Kafka

More jobs like this