← Back to job listings
CO
Senior Specialist Field Engineer (Compute Infrastructure)
CoreWeave · Sunnyvale, United States
About The Role
Join CoreWeave, a leading cloud provider dedicated to powering the AI revolution. As a Senior Specialist Field Engineer, you will be the technical expert for our largest customers, guiding them from facility and rack design to a production-ready supercomputer. You will work closely with various teams, engage hands-on across the customer lifecycle, and shape the future of compute infrastructure.
- Assurer la responsabilité technique pour certains des plus grands clients de l'entreprise, en supervisant le passage de la conception des installations et des racks à un superordinateur prêt pour la production.
- Diriger la mise en service et l'acceptation de nouveaux clusters GPU à grande échelle, en validant le tissu InfiniBand/RoCE et en réalisant des tests de performance HPC.
- Agir en tant que point de contact technique principal pour les clients, en établissant des relations techniques solides et en garantissant leur succès avec les offres d'infrastructure cloud de CoreWeave.
- If you're driven by innovation, thrilled by the possibilities of what specialized compute can enable, and eager to be part of a team that's shaping the future, then CoreWeave is the place for you. Join us and let's embark on this adventure together!
- 7+ years of proven experience as a Solutions Architect, Field Engineer, Infrastructure/Systems Engineer, or Technical Account Manager in Cloud Infrastructure, focusing on building or operating distributed systems or HPC/cloud services, with an expertise focused on bare-metal compute infrastructure and large-scale GPU cluster delivery
- Hands-on experience bringing up, validating, and operating large GPU clusters—including bare metal node pxe boot, hardware health, fabric validation, and HPC acceptance/performance testing—and integrating bare metal with orchestration layers such as Kubernetes and Slurm
- Deep expertise with modern rack-scale GPU server hardware (e.g., NVIDIA HGX / GB200-class systems), high-speed interconnects (InfiniBand, NVLink), and the firmware/BMC/BIOS layer
- Proven track record with building customer relationships, communicating clearly and the ability to break down complex technical concepts to both technical and non-technical audiences
- Expert-level Linux system administration and command-line troubleshooting, paired with strong networking fundamentals (routing, fabric topologies, TCP/IP)
- B.S. in Computer Science or a related technical discipline, or equivalent experience
- You’re curious about the latest and greatest technologies in the AI space
- You’re an expert in managing conflict and achieving mutually beneficial technical outcomes
- You love to help solve challenging technical problems
- Fluency in cloud computing concepts, architecture, and technologies with hands-on experience in designing and implementing cloud solutions
- Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk
- Experience operating security-sensitive, air-gapped, or otherwise locked-down customer environments
- Experience with scripting and automation related to bare-metal provisioning, infrastructure validation, and lifecycle management (Python, Bash, Ansible, or similar)
- Experience designing AI supercomputers from MEP designs
- Experience delivering bare-metal infrastructure at scale for large strategic customers or AI research labs
- Experience with building solutions across multi-cloud or hybrid environment
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring