Skip to content
← Back to job listings

Technical Program Manager (Inference)

CoreWeave · Sunnyvale, United States

External listingfull-timeabout 1 month ago

About The Role

Join CoreWeave's AI/ML TPM team as a Technical Program Manager focused on inference. You will lead complex, cross-functional programs that span inference platform delivery, customer onboarding, launch readiness, and runtime optimization. You will partner with engineering, product, infrastructure, and go-to-market teams to drive programs that improve how inference services are launched, onboarded, operated, and optimized. This role requires strong technical fluency in distributed inference systems, GPU compute, cloud-native architectures, and performance optimization.

  • Drive end-to-end program management for inference platform initiatives, ensuring successful delivery of customer-facing infrastructure and platform capabilities.
  • Lead cross-functional programs for customer onboarding across dedicated and serverless inference offerings, ensuring clear ownership, launch criteria, and readiness for strategic customer use cases.
  • Build and operationalize success metrics, dashboards, launch gates, and review cadences to measure service reliability, onboarding readiness, efficiency, and quality across the inference stack.
  • Strong technical fluency in distributed inference systems, GPU compute, cloud-native architectures, and performance optimization
  • Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk
  • Bachelor's degree in a technical field or equivalent practical experience
  • Experience with inference-serving systems, model onboarding workflows, rollout strategies, and observability tooling
  • Familiarity with launch readiness, supportability, incident follow-through, and release validation for production infrastructure or platform services
  • Demonstrated success driving measurable improvements in reliability, performance, operational readiness, or customer delivery
  • Experience operating in high-growth environments where roadmap execution, reliability expectations, and customer commitments must be managed in parallel
  • Proven experience driving large-scale infrastructure or platform programs from concept to production in complex, cross-functional environments
  • Excellent written and verbal communication skills, with the ability to align engineering, product, infrastructure, and customer-facing stakeholders around shared goals
  • Understanding of customer onboarding for technical products, especially where platform capabilities, infrastructure readiness, and support processes must align for launch
  • 8+ years of technical program management experience in distributed systems, cloud infrastructure, or AI/ML platform engineering
  • You enjoy turning technically complex platform work into predictable execution and successful launches
  • You love driving execution for complex, customer-facing AI infrastructure and platform programs
  • You're effective at creating clarity and momentum across ambiguous, fast-moving multi-team initiatives
  • You're curious about how large-scale inference systems evolve across runtime performance, operational excellence, and customer onboarding

This is an external listing. JobSpring does not represent or verify the employer. Report this listing