← Back to job listings
BR
Senior Software Engineer (Release Infrastructure)
Brex · New York, United States
About The Role
Join Brex as a Senior Software Engineer (Release Infrastructure) and be part of a team that builds systems that scale with speed and intention. You will design, build, and operate the core systems that power Brex's release, observability, and incident management processes. Collaborate with product, platform, and operations teams to ensure safe, fast, and reliable releases, and drive technical strategy and architecture for release and observability systems.
- Design, build, and maintain the release infrastructure that powers Brex’s deployment pipelines and incident workflows.
- Drive technical strategy and architecture for release and observability systems, making them more scalable, reliable, and secure.
- Identify and deliver improvements to the end-to-end release process (from code merge to production) to reduce risk and cycle time.
- Experience architecting and operating scalable, high-availability distributed systems on cloud platforms (e.g., AWS, GCP, Azure)
- Comfort working cross-functionally with product and other engineering teams to debug complex production issues and ship changes safely
- Strong understanding of reliability and SRE practices, including SLIs/SLOs, error budgets, and incident management best practices
- Hands-on experience with CI/CD and release pipelines (e.g., GitHub Actions, CircleCI, Buildkite, Argo, Spinnaker, Jenkins) including build, test, and deployment automation
- Strong communication and collaboration skills, including writing clear design docs and driving technical decisions across teams
- 7+ years of professional experience designing, building, and operating backend or infrastructure systems in production
- Strong proficiency in backend programming languages (e.g., Go, Java, Kotlin, or Python) with a focus on reliability and performance
- Deep familiarity with containerization and orchestration (e.g., Docker, Kubernetes) and infrastructure-as-code (e.g., Terraform, CloudFormation)
- Proven track record of improving release processes (e.g., reducing deployment risk, increasing deployment frequency, automating rollbacks)
- Experience designing and optimizing data storage systems (SQL and/or NoSQL) for operational and observability use cases
- Experience designing and maintaining observability tooling (metrics, logs, tracing) and integrating it into incident response workflows
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring