Skip to content
← Back to job listings

Senior Site Reliability Engineer (SDN)

Lambda · United States

External listingfull-time26 days ago

About The Role

Join Lambda, a leading cloud technology company, as a Senior Site Reliability Engineer. In this role, you will be responsible for building and scaling our cloud offering, operating and improving our multi-tenant cloud networking platform, and developing tooling and automation to improve reliability. You will collaborate with various teams to enhance service reliability and deployment workflows, and drive operational excellence through observability, incident management, and capacity planning. This position offers a flexible work-from-home policy, competitive benefits, and opportunities for professional growth.

  • Operar y escalar la plataforma de red en la nube de Lambda y la infraestructura de SDN.
  • Desarrollar herramientas y automatización para reducir el trabajo operativo y mejorar la confiabilidad.
  • Colaborar con equipos de software, plataforma y redes para mejorar la confiabilidad del servicio y los flujos de trabajo de implementación.
  • Have experience operating and supporting large-scale distributed systems in production
  • Have experience designing and operating CI/CD and GitOps deployment workflows
  • Have experience with observability platforms, monitoring, alerting, and metrics
  • Are comfortable working on the Linux command line and have a solid understanding of the Linux networking stack
  • Have strong troubleshooting skills across Linux systems, Kubernetes, distributed systems, and networking
  • Have experience automating infrastructure and operational workflows using Python, Ansible, or similar tools
  • Have experience with multi-datacenter and hybrid cloud environments
  • Have experience participating in on-call rotations and incident response
  • Have experience with Kubernetes application lifecycle management, upgrades, troubleshooting, and production operations
  • Have 5+ years of experience in Site Reliability Engineering, Production Engineering, or a similar role
  • Experience building and operating Software Defined Networks (SDN), including OpenStack Neutron, OVN, and OVS
  • Experience operating production-scale SDNs in a cloud environment (e.g., infrastructure powering AWS VPC-like networking services)
  • Software development experience in Go and/or Python (C is a plus)
  • Experience automating infrastructure and network configuration using Kubernetes, Helm, Terraform, and Ansible
  • Deep understanding of the Linux networking stack and its interaction with network virtualization technologies, SR-IOV, and DPDK
  • Understanding of the SDN ecosystem and modern cloud networking architectures
  • Experience diagnosing complex production issues across infrastructure, networking, and application layers

This is an external listing. JobSpring does not represent or verify the employer. Report this listing