Summary
WEX is seeking a Site Reliability Engineer 2 to join its Platform Reliability organization supporting the Corporate Payments business. The role focuses on observability, incident response, automation, monitoring, performance optimization, and improving system reliability in collaboration with development teams.
Responsibilities
- Willingness to dig deep into code, networking, operating systems, and/or storage solutions to solve complex issues
- Develop automation and utilize monitoring tools to ensure system reliability
- Participate in incident response and troubleshooting
- Participate in 24x7 Site Reliability rotations and escalation workflows
- Identify and address performance bottlenecks. This will include code optimization, configuration changes, or infrastructure upgrade recommendations
- Collaborate with development teams to ensure software design meets operational requirements
- Continuously improve processes and procedures to increase system reliability and efficiency
- Stay up-to-date with the latest industry trends and technologies
Skills
- 2+ years of hands-on experience as a Site Reliability Engineer or equivalent role
- 2+ years of development experience with at least one major programming language
- Experience with Cloud Computing platforms (AWS, Azure, GCP)
- Ability to thrive in a fast paced, development and operations world
- Strong communication and collaboration skills
- Experience with observability and logging technologies
- Experience with at least one major RDBMS and NoSQL data store
- Experience with containerization technologies such as Docker or Kubernetes
- BA/BS degree in Computer Science or related technical field, or equivalent job experience
- Experience with one or more of the following languages: C#, Java, GoLang, Python
- Experience with infrastructure as code, preferably Terraform
- Working knowledge in building and designing RESTful APIs
- Experience with Grafana and Splunk
- Familiarity with Agile methodologies and practices
- Experience with GitOps
Qualifications
Must Haves
- 2+ years of hands-on experience as a Site Reliability Engineer or equivalent role
- 2+ years of development experience with at least one major programming language
- Experience with Cloud Computing platforms (AWS, Azure, GCP)
- Ability to thrive in a fast paced, development and operations world
- Strong communication and collaboration skills
- Experience with observability and logging technologies
- Experience with at least one major RDBMS and NoSQL data store
- Experience with containerization technologies such as Docker or Kubernetes
- BA/BS degree in Computer Science or related technical field, or equivalent job experience
Nice to Haves
- Experience with one or more of the following languages: C#, Java, GoLang, Python
- Experience with infrastructure as code, preferably Terraform
- Working knowledge in building and designing RESTful APIs
- Experience with Grafana and Splunk
- Familiarity with Agile methodologies and practices
- Experience with GitOps
Benefits
- Non-sales roles are typically eligible for a quarterly or annual bonus based on their role and applicable plan.
- Health insurance
- Dental insurance
- Vision insurance
- Retirement savings plan
- Paid time off
- Health savings account
- Flexible spending accounts
- Life insurance
- Disability insurance
- Tuition reimbursement