Skip to content
← Back to job listings

Software Engineer (Site Reliability Engineering)

Neo4j · London, United Kingdom

External listingfull-time2 months ago

About The Role

Join Neo4j, the leading graph database platform, as a Software Engineer in Site Reliability Engineering (SRE). In this role, you will focus on improving the reliability of Neo4j Aura, our DBaaS product. You will work on building tools, practices, and culture that embed SRE principles at the heart of how Aura operates. Your responsibilities will include automating troubleshooting, treating operations as a software problem, designing for resilience, championing reliability as a product feature, and shaping an observability stack. You will also have the opportunity to collaborate with cross-functional teams, participate in mentorship programs, and engage in employee resource groups.

  • Automate for insight and scale: Build systems that make troubleshooting fast, safe, and scalable across thousands of Neo4j instances.
  • Treat operations as a software problem: Replace tribal knowledge and ad-hoc scripts with tools and systems that codify best practices.
  • Champion reliability as a product feature: Help teams define and act on SLIs and SLOs, turning reliability into a shared, data-driven priority across engineering.
  • Applying SRE practices in real-world environments: defining SLIs and SLOs, reducing toil through automation, and driving reliability through engineering
  • Deploying and managing applications on Kubernetes; cluster-level administration is a plus
  • Designing systems with reliability, safety, and debugability as first-class concerns
  • Writing and contributing to postmortems that lead to meaningful, lasting changes
  • Monitoring distributed systems and understanding their performance characteristics
  • Working with observability tools like OTel Collector, Prometheus, Grafana, and Google Cloud’s operations suite
  • Participating in on-call rotations and incident response with a focus on improvement, not blame
  • Writing backend tools and automation in Go—our primary language—with an emphasis on sound architecture, testing, and maintainability. Strong software skills in other languages, like Python, are also welcome
  • Managing infrastructure with Kustomize and Terraform—keeping it clear, modular, and easy to evolve
  • Troubleshooting large-scale, cloud-based systems with confidence and curiosity
  • Building and maintaining CI/CD workflows—ours run on GitHub Actions
  • Collaborating with other teams to promote SRE thinking—educating on principles like observability, ownership, and service level objectives
  • We're interested in hearing from Engineers with deep experience in some of the following areas:
  • Research shows that members of underrepresented communities are less likely to apply for jobs when they don’t meet all the qualifications. If this is part of the reason you hesitate to apply, we’d encourage you to reconsider and give us the opportunity to review your application

This is an external listing. JobSpring does not represent or verify the employer. Report this listing