← Back to job listings
CL
Senior Site Reliability Engineer (DevEx)
Chainlink Labs · Canada
About The Role
Join Chainlink Labs as a Senior Site Reliability Engineer (DevEx) and play a crucial role in scaling our engineering organization. You will build the infrastructure primitives that define our CI/CD platform, build systems, and developer environments. This is not an operational support role; you will be shaping the reliability and scalability of the developer experience across the entire organization.
- Concevoir et construire les primitives d'infrastructure qui définissent la manière dont notre plateforme CI/CD, nos systèmes de construction et nos environnements de développement évoluent à l'échelle de l'ensemble de l'organisation d'ingénierie.
- Développer les composants d'infrastructure de base, y compris les opérateurs Kubernetes et l'automatisation de mise à l'échelle, que les équipes produit adoptent directement, réduisant ainsi les outils CI/CD et d'environnement sur mesure par équipe.
- Construire l'infrastructure basée sur Kubernetes qui alimente notre plateforme CI/CD, y compris l'infrastructure des runners auto-hébergés GitHub Actions, les applications GitHub et GitHub-as-code.
- AWS/cloud infrastructure production experience
- Hands-on experience with CI/CD platforms at scale: GitHub Actions (self-hosted runners, workflows-as-code), GitHub Apps, and build systems
- Track record of building infrastructure primitives rather than primarily performing support/operations — automation-first mindset
- Proficiency in Go (strongly preferred) or another systems language
- Terraform/GitOps ownership — designing and owning automation, not just running playbooks
- Strong distributed systems and production reliability experience
- 6–9+ years in SRE / Platform / Infrastructure Engineering
- Proven experience scaling Kubernetes in high-throughput production environments
- GitOps workflows (Flux / ArgoCD) experience
- Experience building platform infrastructure, control planes, or Kubernetes Operators (not just consuming them)
- Deep Kubernetes expertise beyond operating clusters — internals, scheduler behavior, custom resources, and cluster-scale failure diagnosis
- Experience with Tailscale or similar zero-trust/overlay networking for CI/CD or remote dev environments
- Experience designing SLO strategies and error-budget usage
- Experience improving diagnosability and observability frameworks (OpenTelemetry or similar)
- Experience building internal developer platforms (IDP) or ephemeral/on-demand dev environments
- Experience building multi-tenant platform infrastructure
- Experience working in high-ambiguity environments
- Experience contributing to Kubernetes ecosystem projects
- Experience with web3 concepts is a plus but not required for this role
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring