← Back to job listings
BA
Engineering Manager (Runtime Fabric)
Baseten · United States
About The Role
Join Baseten as an Engineering Manager for the Runtime Fabrics team. In this role, you will lead the technical direction, grow a world-class team of systems engineers, and shape the infrastructure of Baseten and the open-source container ecosystem. You will oversee the architecture and evolution of the Baseten Delivery Network, drive the expansion of its architecture, and provide technical oversight on GPU-aware isolation mechanisms. You will also champion the team's contributions to the open-source containerd ecosystem and act as the primary advocate for Runtime Fabrics across the organization.
- Lead the Runtime Fabrics team, setting technical direction and ensuring the team's output shapes Baseten's infrastructure and the open-source container ecosystem.
- Recruit, hire, and develop a high-performing team of systems engineers with deep container and Linux expertise, fostering a culture of technical rigor and continuous improvement.
- Oversee the architecture and evolution of the Baseten Delivery Network, driving the expansion of its architecture to include container images, training checkpoints, and deployment artifacts.
- Experience with distributed storage systems, content-addressable storage, or large-scale caching infrastructure
- Strong systems programming background in Go and/or C/C++
- Understanding of how container images are structured, stored, and delivered at scale
- Proven experience managing and growing engineering teams in a systems, infrastructure, or low-level runtime context
- Strong written and verbal communication skills, with the ability to influence without authority across teams
- Contributions to containerd/containerd, opencontainers/runc, google/gvisor, kata-containers/kata-containers, or closely related open-source projects
- Deep familiarity with the Linux container ecosystem: containerd, runc, OCI Runtime Spec, Linux namespaces, and cgroups, with the ability to engage credibly in code reviews and architectural discussions
- Familiarity with lazy-loading snapshotters (stargz, soci, EROFS/Nydus) or peer-to-peer image distribution
- Background in multi-tenant infrastructure or security-sensitive serving environments
- Understanding of containerd's shim API (v2) and experience building custom shim implementations
- Experience with secure container runtimes (gVisor, Sysbox) or micro-VM technologies (Firecracker, Cloud Hypervisor)
- Experience with GPU device access in containers: NVIDIA Container Toolkit, CDI (Container Device Interface), or GPU-aware scheduling
- Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring