Prove logo
Prove
Posted 23 days agoVerified live 10h ago

Site Reliability Engineer

Brief overview

Remote
UndergradOr in progress
$130k–$150k/yrStated range
5+ yrsMinimum
14 H-1B approvalsDept. of Labor
Site Reliability EngineeringAWSKubernetesTerraformOpenTelemetryPrometheusGrafanaJaegerELK StackSplunkGoPython

About the company

Provides digital identity solutions to enterprises without mentioning the company name.

Visa sponsorship history

4 years sponsoring, last filed FY2026

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
14H-1B approved
100%approval rate
3new H-1B hires
$132,000median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
20233
20242
20257
20262
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
20231
20252
20262
Top sponsored roles
Software EngineerSenior Software EngineerSoftware Test Engineer III

Job description

Summary

Prove provides phone-centric identity tokenization and passive cryptographic authentication solutions that support secure digital interactions. The company is seeking a mid- to senior-level Site Reliability Engineer to design, implement, maintain, and deploy highly available, scalable, and reliable systems. The role focuses on observability, AWS infrastructure, automation, infrastructure as code, security, incident response, and service reliability.

Responsibilities

  • Design and implement comprehensive observability solutions across our infrastructure and within applications
  • Establish metrics, logging, and tracing systems that enable quick identification and resolution of issues
  • Create alerting thresholds and automated responses based on service level objectives (SLOs)
  • Provide actionable insights into service to service communications
  • Design, build, and maintain scalable cloud infrastructure on AWS
  • Implement infrastructure-as-code using tools such as Terraform
  • Automate routine operational tasks to reduce toil and improve efficiency
  • Ensure infrastructure security compliance and implement least-privilege access controls
  • Design and implement infrastructure-as-code deployments for container based applications
  • Scale containers based on custom metrics for applications and critical observability infrastructure
  • Conduct thorough post-incident reviews and implement preventative measures
  • Use observability data to perform root cause analysis and system improvements
  • Participate in a 24/7 on call rotation to achieve 99.999% system availability
  • Improve new and existing systems by increasing reliability, performance, and scalability
  • Automate routine operational tasks to reduce toil and improve efficiency
  • Ensure infrastructure security compliance and implement least-privilege access controls
  • Implement efficient infrastructure that balances rapid development and cost
  • Embrace technological changes and development practices while maintaining reliability
  • Participate in a 24/7 on-call rotation
  • Conduct thorough post-incident reviews and implement preventative measures
  • Use observability data to identify system improvements
  • Implement infrastructure as code in a myriad of high compliance development, production, and other environments
  • Scale developer experiences by being the standard bearer of an opinionated platform approach

Skills

  • * 5+ years of experience in Site Reliability Engineering, Platform Engineering or equivalent experience. Software Engineering roles with a strong infrastructure and production engineering aspect also qualify
  • * Expert knowledge of observability platforms and practices (OpenTelemetry, Prometheus, Grafana, Jaeger, ELK stack / Splunk, etc)
  • * Experience with Kubernetes and container orchestration
  • * Strong experience with infrastructure-as-code tools (Terraform, Spacelift, Pulumi)
  • * Proficiency in at least one programming language ( Go, Python )
  • * Deep understanding of cloud platforms, preferably AWS
  • * Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
  • * 3+ years of experience in Site Reliability or Platform Engineering teams
  • * Deep understanding of cloud platforms,  particularly AWS
  • * Strong experience with Kubernetes and container orchestration
  • * Experience withTerraform and  infrastructure-as-code tools
  • * Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
  • * Experience with distributed systems and microservice architectures
  • * Experience working in a high compliance environment
  • * Hand-on experience instrumenting code with OpenTelemetry
  • * Familiarity with service mesh technologies
  • * Contributions to open-source projects
  • * Experience in the identity verification or financial technology industry
  • * Application development experience
  • * Experience with distributed systems and microservice architectures
  • * Experience working in a high compliance environment
  • * Experience with holistic monitoring and alerting for developing platforms
  • * Skilled proficiency in at least one programming language (Go, Python)

Qualifications

Must Haves

  • * 5+ years of experience in Site Reliability Engineering, Platform Engineering or equivalent experience. Software Engineering roles with a strong infrastructure and production engineering aspect also qualify
  • * Expert knowledge of observability platforms and practices (OpenTelemetry, Prometheus, Grafana, Jaeger, ELK stack / Splunk, etc)
  • * Experience with Kubernetes and container orchestration
  • * Strong experience with infrastructure-as-code tools (Terraform, Spacelift, Pulumi)
  • * Proficiency in at least one programming language ( Go, Python )
  • * Deep understanding of cloud platforms, preferably AWS
  • * Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
  • * 3+ years of experience in Site Reliability or Platform Engineering teams
  • * Deep understanding of cloud platforms,  particularly AWS
  • * Strong experience with Kubernetes and container orchestration
  • * Experience withTerraform and  infrastructure-as-code tools
  • * Bachelor's degree in Computer Science, Engineering, or equivalent practical experience

Nice to Haves

  • * Experience with distributed systems and microservice architectures
  • * Experience working in a high compliance environment
  • * Hand-on experience instrumenting code with OpenTelemetry
  • * Familiarity with service mesh technologies
  • * Contributions to open-source projects
  • * Experience in the identity verification or financial technology industry
  • * Application development experience
  • * Experience with distributed systems and microservice architectures
  • * Experience working in a high compliance environment
  • * Experience with holistic monitoring and alerting for developing platforms
  • * Skilled proficiency in at least one programming language (Go, Python)

Benefits

  • Bonus Plan (for eligible roles)
  • Equity Plan
  • Modern Health for financial, mental, and physical wellness
  • 401(k) Retirement Plan & Match (US Offices)
  • Unlimited Vacation and Flexible hours
  • Comprehensive medical benefits for you and your family
  • Emotional & Physical Wellness – Access to wellness services (EAP & Prove Well-Being Reimbursement)
  • Bottomless snacks & beverages for certain office locations
  • Daily GrubHub stipend for lunch if coming into the office (US Offices)

More jobs like this