Allstate logo
Allstate
Posted 3 days agoVerified live 10h ago

Data Engineer (Remote, US)

Brief overview

Remote
UndergradOr in progress
$100k–$170k/yrStated range
4+ yrsMinimum
73 H-1B approvalsDept. of Labor
23 green cardsCertified filings
Data EngineeringApache SparkPythonSQLETL/ELTCloud Data Lakes and LakehousesAnalytical Data ModelingData Quality ManagementData GovernanceData LineageCI/CD PipelinesReal-Time Data Processing

About the company

Leading U.S. insurer providing risk protection and financial services.

Visa sponsorship history

4 years sponsoring, last filed FY2026

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
73H-1B approved
96%approval rate
4new H-1B hires
23PERM certified
$159,503median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
202335
202415
202522
20261
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
20239
20243
20258
Green Card (PERM) FilingsCertified green card filings: a long-term commitment to international hires.
20239
20249
20255
Top sponsored roles
Software Engineer ExpertSenior Software EngineerSoftware Engineer ManagerATSV Security Operations Engineer ExpertLead Software Engineer
Sponsored employees from
IndiaMexicoEgypt

Job description

Summary

Allstate is an insurance company focused on protecting families and their belongings from life’s uncertainties. The Risk Data Team is seeking a Data Engineer to build trusted data products, scalable pipelines, governed datasets, and automation that support risk management, reporting, analytics, and future AI capabilities.

Responsibilities

  • Design, build, and maintain trusted risk data products and scalable batch and streaming data pipelines using cloud-native technologies
  • Integrate data from risk platforms, control systems, vendor tools, operational applications, and external sources into governed, analytics-ready datasets
  • Develop and optimize ETL/ELT workflows that automate data ingestion, transformation, validation, reconciliation, and delivery
  • Build and manage data processing workloads within modern data lake and lakehouse environments, including Microsoft Fabric and OneLake
  • Implement data quality, monitoring, lineage, and reconciliation processes to ensure data reliability, consistency, and accuracy
  • Create curated datasets and analytical outputs that support operational reporting, leadership dashboards, trend analysis, risk identification, and decision-making
  • Optimize data architectures, schemas, and processing patterns for scalability, performance, resilience, and cost efficiency
  • Develop reusable frameworks, engineering standards, CI/CD pipelines, automated testing, and operational monitoring to improve productivity and maintainability
  • Partner with analytics, product, risk, governance, security, and compliance stakeholders to establish data definitions, metrics, quality standards, and secure data access
  • Participate in Agile delivery activities including sprint planning, backlog refinement, design reviews, and continuous improvement initiatives

Skills

  • 4+ years of experience as a Data Engineer or in a similar role building and supporting production-grade data pipelines and data products
  • Hands-on experience with Apache Spark for large-scale data processing and transformation
  • Strong proficiency in Python, SQL, and modern data engineering best practices
  • Experience developing ETL/ELT solutions within cloud-based data lake, lakehouse, or analytics platforms
  • Experience designing and optimizing analytical data models that support reporting, dashboards, operational insights, and advanced analytics
  • Strong understanding of data quality, validation, monitoring, reconciliation, and production support processes
  • Experience with CI/CD pipelines, version control, automated testing, and infrastructure-as-code concepts
  • Demonstrated ability to troubleshoot complex data quality, performance, and integration issues across multiple data sources
  • Strong verbal and written communication skills with the ability to explain technical concepts to both technical and non-technical audiences
  • Experience with Microsoft Fabric, OneLake, or similar modern analytics platforms
  • Experience working with risk, controls, compliance, audit, governance, or regulatory data domains
  • Experience building real-time, near real-time, or event-driven data processing solutions
  • Familiarity with data governance, metadata management, data lineage, access controls, and enterprise data quality frameworks
  • Experience with orchestration and workflow management tools for data pipelines
  • Experience supporting operational reporting, executive dashboards, trend analysis, advanced analytics, or machine learning initiatives
  • Interest in building governed, scalable datasets that support AI and data-driven products
  • Apache Spark, Data Engineering, Data ETL, Python (Programming Language), Software Development, Structured Query Language (SQL)
  • Experience with Microsoft Fabric, OneLake, or similar modern analytics platforms preferred

Qualifications

Must Haves

  • 4+ years of experience as a Data Engineer or in a similar role building and supporting production-grade data pipelines and data products
  • Hands-on experience with Apache Spark for large-scale data processing and transformation
  • Strong proficiency in Python, SQL, and modern data engineering best practices
  • Experience developing ETL/ELT solutions within cloud-based data lake, lakehouse, or analytics platforms
  • Experience designing and optimizing analytical data models that support reporting, dashboards, operational insights, and advanced analytics
  • Strong understanding of data quality, validation, monitoring, reconciliation, and production support processes
  • Experience with CI/CD pipelines, version control, automated testing, and infrastructure-as-code concepts
  • Demonstrated ability to troubleshoot complex data quality, performance, and integration issues across multiple data sources
  • Strong verbal and written communication skills with the ability to explain technical concepts to both technical and non-technical audiences
  • Experience with Microsoft Fabric, OneLake, or similar modern analytics platforms
  • Experience working with risk, controls, compliance, audit, governance, or regulatory data domains
  • Experience building real-time, near real-time, or event-driven data processing solutions
  • Familiarity with data governance, metadata management, data lineage, access controls, and enterprise data quality frameworks
  • Experience with orchestration and workflow management tools for data pipelines
  • Experience supporting operational reporting, executive dashboards, trend analysis, advanced analytics, or machine learning initiatives
  • Interest in building governed, scalable datasets that support AI and data-driven products
  • Apache Spark, Data Engineering, Data ETL, Python (Programming Language), Software Development, Structured Query Language (SQL)

Nice to Haves

  • Experience with Microsoft Fabric, OneLake, or similar modern analytics platforms preferred

Benefits

  • Allstate provides a comprehensive technology setup, including a laptop, monitors, headset, keyboard, and mouse.
  • Employees eligible to work from home also receive a monthly connectivity reimbursement to help offset internet costs.
  • Fully remote work arrangement

More jobs like this