Akamai Technologies logo
Akamai Technologies
Posted 11 days agoVerified live 1d ago

Site Reliability Engineer II

Brief overview

Remote
$95k–$171k/yrStated range
PythonGoShell ScriptingAWSAzureGoogle Cloud PlatformKubernetesLinux/UnixNetworking Protocols (DNS/HTTP/TCP/IP)AnsiblePuppetChefPrometheusGrafanaELK Stack

Job description

Summary

Akamai Technologies builds secure network and cloud infrastructure, including solutions supporting automation, deployment, and monitoring of third-party cloud environments. The Site Reliability Engineer II will provide advanced cloud operations for the Akamai Zero Trust Security platform, ensuring the reliability, performance, availability, and security of critical systems and services.

Responsibilities

  • Ensuring the reliability, availability, and performance of critical systems and services
  • Designing, implementing, and maintaining automation tools and processes to streamline and enhance operational tasks, system provisioning, and configuration management
  • Implementing and maintaining robust monitoring and alerting systems to proactively identify and address potential issues before they impact users
  • Collaborating with the security team to implement and maintain best practices for securing systems and data
  • Identifying opportunities for process improvements, implement best practices, and drive the adoption of new technologies to enhance system reliability and efficiency
  • Identifying and optimize performance bottlenecks in systems, applications, and infrastructure components
  • Maintaining high availability of an infrastructure. Continuous integration and continuous delivery of the platform

Skills

  • Have 2 years of relevant experience and a Bachelor's degree in Computer Science or its equivalent
  • Possess experience as a Site Reliability Engineer or in a similar role
  • Have programming and scripting skills (e.g., Python, Go, Shell)
  • Have experience with cloud platforms (e.g., AWS, Azure, GCP) and container orchestration tools (e.g., Kubernetes)
  • Possess in-depth knowledge of Linux/Unix systems and networking concepts, including network protocols such as DNS/HTTP/TCP/IP
  • Have proficiency in configuration management tools (e.g., Ansible, Puppet, Chef)
  • Possess familiarity with monitoring and logging tools (e.g., Prometheus, Grafana, ELK stack)
  • Able to work in a dynamic environment
  • Must have good problem solving skills
  • Familiarity with DevOps principles and practices is ideal

Qualifications

Must Haves

  • Have 2 years of relevant experience and a Bachelor's degree in Computer Science or its equivalent
  • Possess experience as a Site Reliability Engineer or in a similar role
  • Have programming and scripting skills (e.g., Python, Go, Shell)
  • Have experience with cloud platforms (e.g., AWS, Azure, GCP) and container orchestration tools (e.g., Kubernetes)
  • Possess in-depth knowledge of Linux/Unix systems and networking concepts, including network protocols such as DNS/HTTP/TCP/IP
  • Have proficiency in configuration management tools (e.g., Ansible, Puppet, Chef)
  • Possess familiarity with monitoring and logging tools (e.g., Prometheus, Grafana, ELK stack)
  • Able to work in a dynamic environment
  • Must have good problem solving skills

Nice to Haves

  • Familiarity with DevOps principles and practices is ideal

Benefits

  • Remote work, office work, or a combination of both through Akamai's FlexBase program
  • Annual bonus or incentives
  • Equity awards
  • Employee Stock Purchase Plan (ESPP)
  • Healthcare
  • 401K savings plan
  • Company holidays
  • Vacation in the form of PTO
  • Sick time
  • Parental leave
  • Employee assistance program including a focus on mental and financial wellness

More jobs like this