Skip to content
← Back to job listings

Senior Site Reliability Engineer

Talkiatry · United States

External listingfull-time14 days ago

About The Role

Join Talkiatry, a leading telehealth company, as our first Senior Site Reliability Engineer. In this high-leverage, founding role, you will define and implement SRE principles across our six product teams, making reliability a shared responsibility. Your success will be measured by reducing production outages, improving incident detection, and enhancing observability. Enjoy benefits such as a 401(k) match, paid parental leave, and health insurance coverage from day one.

  • Defining and rolling out an SRE practice for a six-team organization, including SLOs/SLIs, error budgets, and reliability standards.
  • Building and improving observability, including metrics, logging, distributed tracing, dashboards, and alerting, to ensure incidents are detected by monitoring.
  • Driving down outage frequency by surfacing systemic reliability risks and partnering with teams to remediate them at the root.
  • Proficiency with infrastructure-as-code (e.g., Terraform) and comfort building automation and tooling (Python, TypeScript, or similar)
  • A track record of reducing incidents and improving detection—the outcomes this role is judged on
  • Strong experience operating production systems on AWS
  • Deep observability expertise across metrics, logging, tracing, and alerting (e.g., Datadog, Prometheus, Grafana, or similar)
  • Excellent communication skills, with the ability to influence and align teams you don't directly manage
  • Hands-on experience defining SLOs/SLIs and using error budgets to guide engineering decisions
  • 7+ years in software or infrastructure engineering, with substantial hands-on SRE or production reliability experience
  • Experience standing up an SRE function for the first time at a startup or scale-up
  • Familiarity with the stack our teams run on (TypeScript/Node.js, React, AWS EKS & RDS)
  • Background in healthcare, tele-health, or other regulated, compliance-sensitive environments (e.g., HIPAA)
  • Kubernetes or container orchestration experience

This is an external listing. JobSpring does not represent or verify the employer. Report this listing