Skip to content
← Back to job listings

Site Reliability Engineering Tech Lead

DataHub · Palo Alto, United States

External listingfull-time24 days ago

About The Role

Join DataHub as a Site Reliability Engineering Tech Lead, where you'll drive the reliability, scalability, and operational excellence of our platform offerings. You'll lead technical initiatives across DataHub Cloud and our enterprise deployment solution, design and implement robust infrastructure solutions, and collaborate with cross-functional teams to ensure reliable product delivery. This remote-friendly position offers a range of benefits, including parental leave, medical, dental, and vision insurance, equity, and a monthly co-working space budget.

  • Design and implement robust, scalable infrastructure solutions for DataHub Cloud and enterprise deployments, leading the technical vision for multi-cloud deployment strategies.
  • Drive best practices for infrastructure as code, configuration management, and deployment automation, and partner with product and engineering teams to influence the development of advanced deployment capabilities.
  • Lead incident response and post-mortem processes to drive continuous improvement, implement chaos engineering practices to proactively identify system weaknesses, and optimize system performance, capacity planning, and cost efficiency.
  • 8+ years of experience in Site Reliability Engineering, Platform Engineering, or DevOps roles
  • Proficiency in containerization technologies (Docker, Kubernetes) and orchestration
  • Experience with CI/CD pipelines and deployment automation
  • Strong knowledge of networking, security, and database operations in cloud environments
  • Strong programming skills in Python, Java, or similar languages
  • Strong expertise with cloud platforms (AWS, GCP, Azure) and infrastructure automation tools
  • 3+ years of technical leadership experience managing engineering teams
  • Deep understanding of monitoring and observability tools (Prometheus, Grafana, Datadog, etc.)
  • Experience with infrastructure as code tools (Terraform, CloudFormation, Pulumi)
  • Experience building and operating multi-tenant SaaS platforms
  • Knowledge of data infrastructure and metadata management systems
  • Background in developing customer-facing deployment and management tools
  • Previous experience in a customer-facing technical role or working with enterprise clients
  • Experience with service mesh technologies and microservices architectures
  • Experience with data governance or data catalog platforms

This is an external listing. JobSpring does not represent or verify the employer. Report this listing