Skip to content
← Back to job listings

Senior Machine Learning Engineer (Large Systems)

Graphcore · Cambridge, United Kingdom

External listingfull-time28 days ago

About The Role

Join Graphcore, a leading AI technology company, as a Senior Machine Learning Engineer. In this role, you will develop and optimize AI models for our specialized hardware, work on large-scale systems, and collaborate with various teams to advance AI technology. Enjoy a flexible work-life balance, private medical insurance, a pension plan, and opportunities for career progression and personal development.

  • Contribuer au développement et à l'optimisation des modèles d'IA adaptés au matériel spécialisé de Graphcore.
  • Travailler sur des systèmes à grande échelle où la performance est critique, en collaboration avec les équipes de développement logiciel et de recherche.
  • Participer à la construction d'applications de référence, à l'optimisation des bibliothèques logicielles clés et à la collaboration avec l'équipe de recherche.
  • We seek engineers with strong technical skills and an understanding of AI model implementation at scale, eager to make a tangible impact in this rapidly evolving field
  • If you're excited about advancing the next generation of AI models on cutting-edge hardware, we’d love to hear from you!
  • Ability to move quickly in a dynamic environment
  • Strong communicator - able to explain complex technical concepts to different audiences
  • Developed deep understanding of performance bottlenecks and how to overcome them
  • Enjoy cross-functional work collaborating with other teams
  • Experience in distributed training or inference of ML models across 64+ accelerators
  • Expertise in deep learning from model training to optimisation and evaluation
  • Proficiency in deep learning frameworks like PyTorch/JAX
  • Capable of designing, executing and reporting from ML experiments
  • Bachelor/Master's/PhD or equivalent experience in Machine Learning, Computer Science, Maths, Data Science, or related field
  • Strong Python or C++ software development skills
  • Knowledge of cloud computing platforms
  • Familiarity with HPC systems and networking including Infiniband, NVLink, RoCE technologies
  • Have contributed to open-source projects or published research papers in relevant fields
  • Keen to present, publish and deliver talks in the AI community
  • Efficient computing based on low-precision arithmetic
  • Building production systems with large language models
  • Experience writing C++/Triton/CUDA kernels for performance optimisation of ML models
  • MLOps for Kubernetes-based clusters
  • Experience in one or more of:

This is an external listing. JobSpring does not represent or verify the employer. Report this listing