Site Reliability Engineering (Fabric)
MongoDB · United States
About The Role
Join MongoDB's Site Reliability Engineering (SRE) team as a Site Reliability Engineer (SRE) on the Fabric team. In this role, you will be responsible for building and maintaining the infrastructure necessary for secure and efficient communication between our services. You will collaborate with service-owning teams, participate in a 24/7 on-call rotation, and leverage your expertise in networking, distributed systems, and automation. The ideal candidate will have 10+ years of experience in software and distributed systems, a strong knowledge of service mesh and load-balancing concepts, and familiarity with modern cloud-based infrastructure.
- Participate in the development of a reliable and resilient multi-cloud globally-connected network that is crucial for MongoDB’s services.
- Collaborate with service-owning teams to provide internal support, addressing technical issues and offering guidance on best practices for service-to-service connectivity.
- Participate in a 24/7 on-call rotation to swiftly resolve issues related to network architecture and service-to-service connectivity, ensuring minimal disruption and high availability.
- We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team
- Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles
- Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”)
- Possess a customer-focused mindset, driving improvements that benefit end-users
- Have a strong knowledge of service mesh and load-balancing concepts, and be eager to implement these in a multi-cloud environment
- Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring