← Back to job listings
SA
AI Infrastructure Engineer (Model Serving Platform)
Scale AI · New York, United States
About The Role
Join our ML Infrastructure team as an AI Infrastructure Engineer. In this role, you will design and build platforms for scalable and efficient serving of LLMs and AI agents. You will collaborate with researchers and engineers, conduct architecture and design reviews, develop monitoring solutions, and lead projects from start to finish. Enjoy comprehensive health coverage, personal and career growth opportunities, a supportive community, and parental support.
- Design and build platforms for scalable, reliable, and efficient serving of LLMs and AI agents.
- Collaborate with researchers and engineers to integrate and optimize models for production and research use cases.
- Lead projects end-to-end, from requirements gathering to implementation, in a cross-functional environment.
- The ideal candidate combines strong ML fundamentals with deep expertise in backend system design
- Familiarity with cloud infrastructure (AWS, GCP) and infrastructure as code (e.g., Terraform)
- Proven ability to solve complex problems and work independently in fast-moving environments
- Deep understanding of concurrency, memory management, networking, and distributed systems
- Experience with containers, virtualization, and orchestration tools (e.g., Docker, Kubernetes)
- Strong programming skills in one or more languages (e.g., Python, Go, Rust, C++)
- Experience with modern LLM serving frameworks such as vLLM, SGLang, TensorRT-LLM, or text-generation-inference
- 4+ years of experience building large-scale, high-performance backend systems
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring