← Back to job listings
MA
Site Reliability Engineer
Mistral Ai · New York, New York, United States
About The Role
Join Mistral, a leading provider of full-stack AI solutions. As a Site Reliability Engineer, you will play a crucial role in shaping the reliability, scalability, and performance of our platform and customer-facing applications. You will work closely with software engineers and research teams to ensure our systems meet and exceed customer expectations. Your responsibilities will include designing and maintaining scalable infrastructures, troubleshooting production issues, implementing monitoring and incident response systems, and driving continuous improvement in infrastructure automation.
- Design, build, and maintain scalable, highly available, and fault-tolerant infrastructures to support web services and ML workloads.
- Implement and improve monitoring, alerting, and incident response systems to ensure optimal system performance and minimize downtime.
- Drive continuous improvement in infrastructure automation, deployment, and orchestration using tools like Kubernetes, Flux, Terraform.
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring