Trupanion logo
Trupanion
Posted 54 days agoVerified live 17h ago

Data Engineer

Brief overview

Remote
UndergradOr in progress
$130k–$150k/yrStated range
2+ yrsMinimum
DatabricksMicrosoft AzureDelta LakeApache SparkPySparkSparkSQLUnity CatalogDelta Live TablesETL/ELT PipelinesData MartsData Quality Testing and MonitoringCI/CD Pipelines

About the company

Trupanion logo
Trupaniontrupanion.com

Trupanion is an insurance company which provides pet care insurance.

Job description

Summary

Trupanion is a leading provider of medical insurance for cats and dogs in North America, offering a collaborative, casual, and pet-friendly work environment. The Data Engineer will design, build, optimize, and document reliable data systems and data marts on a Databricks Lakehouse platform hosted in Microsoft Azure, while improving data quality, governance, pipeline reliability, and collaboration with analytics and data science teams.

Responsibilities

  • Data Ingestion & Pipeline Reliability: Responsible for ensuring continuous, accurate data availability by monitoring, troubleshooting, and optimizing batch and real-time streaming pipelines across the Medallion architecture to deliver stable, high-availability data feeds with minimal downtime
  • Data Mart Delivery & Business Modeling: Responsible for delivering query-ready Gold-layer data marts, dimensional models, and aggregate tables by partnering directly with Data Analytics and Data Science teams to map business requirements into high-performance user-facing schemas
  • Data Contract Validation & Ingestion Quality: Responsible for protecting the integrity of the Lakehouse by configuring automated validation gates and alerts at the ingestion stage to catch and block corrupt or non-compliant source data
  • Data Quality Frameworks & Precision: Responsible for maintaining flawless data integrity and trust across all data products by writing, deploying, and maintaining automated data quality test suites and expectations for key transformation stages
  • Data Governance & Asset Discoverability: Responsible for ensuring all data assets are easily discoverable by documenting schemas, applying metadata tags
  • Cross-Functional Collaboration & Autonomy: Responsible for serving as an autonomous data engineering partner to downstream teams by proactively investigating data discrepancies, resolving support tickets, and aligning business definitions to unblock analysts and data scientists

Skills

  • Bachelor's level degree or higher in Computer Science, Engineering, or equivalent practical experience
  • Minimum 2+ years of experience designing, building, and enhancing high-performance pipelines (ETL/ELT) and building optimized Data Marts across the enterprise
  • Minimum 2+ years of deep technical experience within a Databricks-centric environment leveraging Delta Lake & Apache Spark (PySpark / SparkSQL), Unity Catalog for data governance, lineage, and documentation, Delta Live Tables (DLT) or similar orchestration tools
  • Proven experience implementing robust data quality, testing, and monitoring frameworks (e.g., Great Expectations, Databricks Expectations, or custom data quality gates) to ensure flawless data integrity
  • Experience required with software and infrastructure change management, CI/CD pipelines (Azure DevOps/GitHub Actions), and source code control
  • Passion for the latest industry trends, including the Medallion Architecture (Bronze/Silver/Gold), Data Contracts, and decentralized data design
  • Strong conceptual thinker who doesn't just build, but actively designs and documents data assets, schemas, and definitions directly within Unity Catalog
  • Hands-on experience implementing data pipelines for a variety of flows, including batch data integrations, real-time streaming analytics (Structured Streaming), and big data analytics
  • Experience setting up comprehensive logging, alerting, and monitoring solutions for Spark jobs and data pipelines
  • Thrives in a fast-paced, Agile software development environment, with a strong proactive drive to learn and adopt new tools and features as the Databricks ecosystem evolves

Qualifications

Must Haves

  • Bachelor's level degree or higher in Computer Science, Engineering, or equivalent practical experience
  • Minimum 2+ years of experience designing, building, and enhancing high-performance pipelines (ETL/ELT) and building optimized Data Marts across the enterprise
  • Minimum 2+ years of deep technical experience within a Databricks-centric environment leveraging Delta Lake & Apache Spark (PySpark / SparkSQL), Unity Catalog for data governance, lineage, and documentation, Delta Live Tables (DLT) or similar orchestration tools
  • Proven experience implementing robust data quality, testing, and monitoring frameworks (e.g., Great Expectations, Databricks Expectations, or custom data quality gates) to ensure flawless data integrity
  • Experience required with software and infrastructure change management, CI/CD pipelines (Azure DevOps/GitHub Actions), and source code control
  • Passion for the latest industry trends, including the Medallion Architecture (Bronze/Silver/Gold), Data Contracts, and decentralized data design
  • Strong conceptual thinker who doesn't just build, but actively designs and documents data assets, schemas, and definitions directly within Unity Catalog
  • Hands-on experience implementing data pipelines for a variety of flows, including batch data integrations, real-time streaming analytics (Structured Streaming), and big data analytics
  • Experience setting up comprehensive logging, alerting, and monitoring solutions for Spark jobs and data pipelines
  • Thrives in a fast-paced, Agile software development environment, with a strong proactive drive to learn and adopt new tools and features as the Databricks ecosystem evolves

Benefits

  • Monthly bonuses
  • Restricted Stock Units granted to all new team members; new hire grants vest over 4 years
  • Full medical, dental, and vision benefits at no cost to the employee
  • Four weeks of paid time off and 9 paid float holidays
  • Five-week sabbatical after five years of employment
  • Open, casual, pet-friendly, and fun office environment
  • Free medical health insurance for your pet (1 dog or cat)
  • Paid time off to volunteer at nonprofit organizations
  • Seattle candidates will have a hybrid remote/in-office schedule
  • Free on-site gym
  • Free dog walking services for office pets during business hours
  • Free parking
  • Paid ORCA cards

More jobs like this