Summary
Semarchy is seeking a Junior Site Reliability Engineer to support the reliability, scalability, and availability of its SaaS products and infrastructure. The role involves cloud infrastructure automation, monitoring, incident response, security maintenance, performance optimization, and cost management while working with senior engineers and support teams.
Responsibilities
- Help build and maintain systems that support high availability, fault tolerance and disaster recovery across our cloud environments
- Track SLOs and SLIs for critical services, and help improve them over time
- Write and maintain automation scripts and Infrastructure as Code (e.g., Terraform) to cut down on manual operational work
- Set up and tune monitoring, alerting and logging so issues get caught early
- Follow security and change-management practices that support our SOC 2 and ISO 27001 commitments
- Help remediate vulnerabilities and keep infrastructure patched and hardened
- Take part in on-call rotations, with support from senior team members, to respond to incidents and outages
- Contribute to post-incident reviews and follow through on the actions that prevent repeat issues
- Write and keep up-to-date documentation and runbooks
- Monitor system performance, capacity and cloud spend, and flag areas to optimize
- Help carry out capacity upgrades and performance improvements
Skills
- • 2–3 years in an SRE, DevOps, Cloud Infrastructure or similar role
- • Hands-on experience with AWS (e.g., EC2, EKS, RDS, S3)
- • Working knowledge of Kubernetes and containers (Docker, Helm a plus)
- • Some experience with an IaC tool such as Terraform, CloudFormation or Pulumi
- • Familiarity with monitoring, alerting and logging tools (e.g., Datadog, Prometheus, Grafana, CloudWatch)
- • Scripting in Python, Go or Bash
- • Good troubleshooting skills and eagerness to learn from incidents
- • Bachelor's degree in Computer Science, Information Technology or a related field (or equivalent experience)
- • Experience with CI/CD pipelines (e.g., GitHub Actions, GitLab CI, ArgoCD)
- • Experience with cloud security monitoring tools (e.g., AWS GuardDuty, Wiz)
- • Familiarity with Agile, DevOps and DevSecOps practices
Qualifications
Must Haves
- • 2–3 years in an SRE, DevOps, Cloud Infrastructure or similar role
- • Hands-on experience with AWS (e.g., EC2, EKS, RDS, S3)
- • Working knowledge of Kubernetes and containers (Docker, Helm a plus)
- • Some experience with an IaC tool such as Terraform, CloudFormation or Pulumi
- • Familiarity with monitoring, alerting and logging tools (e.g., Datadog, Prometheus, Grafana, CloudWatch)
- • Scripting in Python, Go or Bash
- • Good troubleshooting skills and eagerness to learn from incidents
Nice to Haves
- • Bachelor's degree in Computer Science, Information Technology or a related field (or equivalent experience)
- • Experience with CI/CD pipelines (e.g., GitHub Actions, GitLab CI, ArgoCD)
- • Experience with cloud security monitoring tools (e.g., AWS GuardDuty, Wiz)
- • Familiarity with Agile, DevOps and DevSecOps practices