Summary
MongoDB is a leading data platform company that empowers innovation. They are seeking a Site Reliability Engineer to join the Fabric team, responsible for building and maintaining infrastructure for secure communication between services.
Responsibilities
- Participate in the development of a reliable and resilient multi-cloud globally-connected network that is crucial for MongoDB’s services
- Collaborate with service-owning teams to provide internal support, addressing technical issues and offering guidance on best practices for service-to-service connectivity
- Participate in a 24/7 on-call rotation to swiftly resolve issues related to network architecture and service-to-service connectivity, ensuring minimal disruption and high availability
Skills
- Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles
- Possess a customer-focused mindset, driving improvements that benefit end-users
- Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”)
- Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs
- Have a strong knowledge of service mesh and load-balancing concepts, and be eager to implement these in a multi-cloud environment
Qualifications
Must Haves
- Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles
- Possess a customer-focused mindset, driving improvements that benefit end-users
- Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”)
- Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs
- Have a strong knowledge of service mesh and load-balancing concepts, and be eager to implement these in a multi-cloud environment
Benefits
- Equity
- Participation in the employee stock purchase program
- Flexible paid time off
- 20 weeks fully-paid gender-neutral parental leave
- Fertility and adoption assistance
- 401(k) plan
- Mental health counseling
- Access to transgender-inclusive health insurance coverage
- Health benefits offerings