Director of Platform & Reliability Engineering
Forge · New York, United States
About The Role
Join Forge as the Director of Platform & Reliability Engineering, where you will lead a critical engineering organization responsible for the systems, services, and operational practices that enable Forge to build and run secure, scalable, and highly reliable products. You will oversee Platform Engineering, Cloud Operations, and Site Reliability Engineering, setting the vision for how internal platforms, cloud infrastructure, developer enablement, and production operations evolve to support the company's growth. This is an exciting opportunity for a seasoned technical and people leader who can operate strategically while remaining close enough to architecture, delivery, and operations to guide strong technical decision-making.
- Lead and develop the Platform Engineering, Cloud Engineering, and Site Reliability Engineering teams, including organizational design, hiring, coaching, and performance management.
- Define and execute the strategy for internal platforms, cloud infrastructure, reliability engineering, observability, and developer enablement.
- Drive improvements in availability, performance, scalability, security, and operational maturity across production systems.
- Experience defining engineering strategy, driving cross-functional alignment, and translating business priorities into platform and infrastructure roadmaps
- 8+ years of software engineering experience, including significant time leading infrastructure, platform, cloud, or reliability-focused teams
- Deep experience with cloud infrastructure, infrastructure as code, observability, incident response, and modern platform engineering practices
- Excellent communication and stakeholder management skills, with the ability to influence technical and non-technical leaders
- Bachelor's degree in Computer Science or a closely related field, or equivalent practical experience
- Experience in FinTech, financial services, or another regulated industry
- 5+ years of people leadership experience, including leading managers and building high-performing engineering organizations
- Strong technical judgment in distributed systems, production operations, service reliability, and scalable engineering architecture
- Strong familiarity with Kubernetes, container platforms, CI/CD systems, and infrastructure automation tooling
- Experience leading organizations through cloud modernization, platform standardization, or large-scale reliability transformations
- Experience building developer platforms and self-service infrastructure capabilities that improve engineering productivity
- Experience at growth-stage companies where balancing scale, speed, and reliability is essential
- Physical requirements: operate a computer for 8 hours per day; give and receive detailed information through verbal and written communication
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring