Skip to content
← Back to job listings

Senior Software Engineer (Site Reliability)

Parabola · United States

External listingfull-time22 days ago

About The Role

Join our rapidly growing team as a Senior Site Reliability Engineer. In this role, you will measure and improve the performance of our software, extend and monitor our infrastructure stack, and collaborate with our engineering team and company leadership. You will also work on our core orchestration logic, optimize our services for scalability and stability, and improve our developer experience. This position offers excellent benefits, including medical, dental, and vision coverage, parental leave, remote and hybrid work opportunities, and unlimited PTO.

  • Measuring software performance and driving improvements based on established baselines.
  • Monitoring core business-logic software and extending the infrastructure stack to ensure reliability.
  • Collaborating with the engineering team and company leadership to enhance site reliability and optimize services.
  • You’re not afraid to ask for help, and you’re happy to give it, too
  • Familiarity with containerization and orchestration tools (e.g., Docker, Kubernetes) and how they interact with backend services, and with Linux
  • Experience implementing and managing AWS infrastructure
  • You're excited to join a hybrid team and work out of our NYC or SF office ~3 days a week
  • A proven record of building efficient, performant, and easy to extend systems
  • Has maintained quantitative metrics of site reliability, while also demonstrating judgment about appropriate strictness for SLOs & SLAs. Given our team size, the expectations are somewhat less formal and mature than many SRE teams have; you will both strengthen our approach, but also navigate tradeoffs
  • 5+ years of SRE, DevOps, or Platform engineering experience
  • You’re an enthusiastic communicator and you like working with a team that provides both mutual support and thoughtful critique
  • Experience building AI platforms, or the tooling and systems behind AI products
  • Experience with Temporal
  • Familiarity with Kubernetes/Helm (we use Amazon EKS)
  • Experience managing CI/CD pipeline for multiple environments, and developer experience more broadly
  • Experience building cloud storage or data modeling products
  • Experience implementing and deploying infrastructure-as-a-service (IaaS) tools to manage production cloud environments
  • Experience working at early stage startups

This is an external listing. JobSpring does not represent or verify the employer. Report this listing