← Back to job listings
OK
Senior Software Engineer (Observability)
Okta · Washington, United States
About The Role
Join the Auth0 Platform Observability team as a Senior Software Engineer. In this role, you will be a core technical leader, working cross-functionally to integrate services with instrumentation libraries, support product teams, and investigate incidents to identify observability gaps. You will also design, build, and maintain scalable observability infrastructure, troubleshoot performance issues, and automate operational tasks. This position offers work-from-home opportunities, health and wellness benefits, financial benefits, and time off.
- Act as a core technical leader on the Observability team, working cross-functionally to integrate services with instrumentation libraries and support product teams.
- Champion observability best practices, educating and correcting anti-patterns, and teaching other engineering teams how to build robust instrumentation.
- Design, build, and maintain scalable observability infrastructure using tools like Terraform, and troubleshoot performance and operational issues.
- We are looking for engineers passionate about monitoring, observing, measuring uptime and availability, and ensuring platform stability. If you have experience within the Site Reliability Engineering (SRE) field or working as a Development Operations (DevOps) engineer, and you have a passion for Observability tooling, this position will allow you to further your learning and development in these areas
- Proven ability to lead cross-functional technical initiatives and collaborate seamlessly with multiple engineering teams
- Deep understanding of microservice architecture and best practices
- Strong coding skills in Node.js or Golang
- 5+ years of platform engineering, SRE, or DevOps experience
- Experience with cloud infrastructure like AWS, Google Cloud, or Azure
- A data-driven approach to debugging complex, cross-service performance bottlenecks
- Expertise in the Datadog ecosystem (Metrics, Logs, Traces, and Error Tracking), including establishing alerting standards, implementing tagging taxonomies, and managing Datadog configurations via Terraform
- Experience with containerization and orchestration tools (e.g., Docker, Kubernetes)
- Experience in coaching and mentoring more junior engineers
- Hands-on experience with OpenTelemetry (OTel), Vector, or similar frameworks for instrumenting applications
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring