Kohls logo
Kohls
Posted 47 days agoVerified live 2d ago

Reliability Engineer (Remote)

Brief overview

Remote
UndergradOr in progress
2+ yrsMinimum
84 H-1B approvalsDept. of Labor
28 green cardsCertified filings
JavaPythonGoNode.jsSystems ArchitectureOperating System InternalsCloud ComputingCloudWatchGrafanaPrometheusOpenTelemetryDockerKubernetesRancherNetwork Fundamentals

About the company

A high-profile retailer focused on caring for families and building a healthy, safe workplace for associates.

Visa sponsorship history

4 years sponsoring, last filed FY2026

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
84H-1B approved
99%approval rate
8new H-1B hires
28PERM certified
$155,000median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
202326
202416
202541
20261
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
20236
202411
20253
20267
Green Card (PERM) FilingsCertified green card filings: a long-term commitment to international hires.
202318
20243
20255
20262
Top sponsored roles
Senior Software EngineerStaff Software EngineerManager, Advanced Analytics and ResearchSenior Reliability EngineerStaff Reliability Engineer
Sponsored employees from
IndiaDenmark

Job description

Summary

Kohl's is seeking a Reliability Engineer to ensure the resilience and availability of its systems and applications. The role focuses on incident response, root cause analysis, automation, observability, capacity planning, chaos engineering, and reliability best practices across product teams.

Responsibilities

  • Drive incident response efforts, perform root cause analysis and implement preventative measures to enhance system reliability
  • Establish consistent practices that elevate Kohl’s operational excellence through automation and process improvements
  • Follow software lifecycle and drive reliability, observability and efficiency across product teams within an assigned domain
  • Identify repeated toil and find opportunities for automation and risk reduction
  • On-call on a rotation to respond to production incidents and conduct blameless retros and root-cause analyses (RCAs) to drive a culture of continuous improvements
  • Proactively identify failures before they cause outages using chaos engineering techniques such as edge cases, failure modes and design review
  • Advise on capacity planning and provide continuous assessments on systems behavior and consumption
  • Work with product managers to identify and prioritize work for reliability best practices (i.e., leveraging SLIs/SLOs/Error Budgets)
  • Additional tasks may be assigned

Skills

  • Bachelor's Degree or equivalent in MIS, Computer Science or related field
  • 2+ years of experience in software development
  • Strong programming skills in one or more languages (Java, Python, Go or Node.js)
  • Working knowledge of systems architecture, operating system internals and network fundamentals
  • Experience working with one cloud platform (e.g., GCP, AWS, or Azure)
  • Experience with monitoring techniques and tools (e.g., CloudWatch, Grafana, Prometheus, OpenTelemetry, Tracing)
  • Working knowledge around containerization and container orchestration (e.g., Docker, Kubernetes, Rancher)

Qualifications

Must Haves

  • Bachelor's Degree or equivalent in MIS, Computer Science or related field
  • 2+ years of experience in software development
  • Strong programming skills in one or more languages (Java, Python, Go or Node.js)
  • Working knowledge of systems architecture, operating system internals and network fundamentals
  • Experience working with one cloud platform (e.g., GCP, AWS, or Azure)

Nice to Haves

  • Experience with monitoring techniques and tools (e.g., CloudWatch, Grafana, Prometheus, OpenTelemetry, Tracing)
  • Working knowledge around containerization and container orchestration (e.g., Docker, Kubernetes, Rancher)

More jobs like this