Skip to content
← Back to job listings

Senior Specialist Field Engineer (Compute Infrastructure)

CoreWeave · Sunnyvale, United States

External listingfull-timeabout 2 months ago

About The Role

Join CoreWeave, a leading cloud provider dedicated to powering the AI revolution. As a Senior Specialist Field Engineer, you will be the technical expert for our largest customers, guiding them from facility and rack design to a production-ready supercomputer. You will work closely with various teams, engage hands-on across the customer lifecycle, and shape the future of compute infrastructure.

  • Assurer la responsabilité technique pour certains des plus grands clients de l'entreprise, en supervisant le passage de la conception des installations et des racks à un superordinateur prêt pour la production.
  • Diriger la mise en service et l'acceptation de nouveaux clusters GPU à grande échelle, en validant le tissu InfiniBand/RoCE et en réalisant des tests de performance HPC.
  • Agir en tant que point de contact technique principal pour les clients, en établissant des relations techniques solides et en garantissant leur succès avec les offres d'infrastructure cloud de CoreWeave.
  • If you're driven by innovation, thrilled by the possibilities of what specialized compute can enable, and eager to be part of a team that's shaping the future, then CoreWeave is the place for you. Join us and let's embark on this adventure together!
  • 7+ years of proven experience as a Solutions Architect, Field Engineer, Infrastructure/Systems Engineer, or Technical Account Manager in Cloud Infrastructure, focusing on building or operating distributed systems or HPC/cloud services, with an expertise focused on bare-metal compute infrastructure and large-scale GPU cluster delivery
  • Hands-on experience bringing up, validating, and operating large GPU clusters—including bare metal node pxe boot, hardware health, fabric validation, and HPC acceptance/performance testing—and integrating bare metal with orchestration layers such as Kubernetes and Slurm
  • Deep expertise with modern rack-scale GPU server hardware (e.g., NVIDIA HGX / GB200-class systems), high-speed interconnects (InfiniBand, NVLink), and the firmware/BMC/BIOS layer
  • Proven track record with building customer relationships, communicating clearly and the ability to break down complex technical concepts to both technical and non-technical audiences
  • Expert-level Linux system administration and command-line troubleshooting, paired with strong networking fundamentals (routing, fabric topologies, TCP/IP)
  • B.S. in Computer Science or a related technical discipline, or equivalent experience
  • You’re curious about the latest and greatest technologies in the AI space
  • You’re an expert in managing conflict and achieving mutually beneficial technical outcomes
  • You love to help solve challenging technical problems
  • Fluency in cloud computing concepts, architecture, and technologies with hands-on experience in designing and implementing cloud solutions
  • Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk
  • Experience operating security-sensitive, air-gapped, or otherwise locked-down customer environments
  • Experience with scripting and automation related to bare-metal provisioning, infrastructure validation, and lifecycle management (Python, Bash, Ansible, or similar)
  • Experience designing AI supercomputers from MEP designs
  • Experience delivering bare-metal infrastructure at scale for large strategic customers or AI research labs
  • Experience with building solutions across multi-cloud or hybrid environment

This is an external listing. JobSpring does not represent or verify the employer. Report this listing