← Back to job listings
GR
Senior Machine Learning Engineer (Large Systems)
Graphcore · London, United Kingdom
About The Role
Join Graphcore, a leading AI technology company, as a Senior Machine Learning Engineer. In this role, you will develop and optimize AI models for our specialized hardware, work on large-scale systems, and collaborate with various teams to advance AI technology. Enjoy a flexible work-life balance, private medical insurance, a pension plan, and opportunities for career progression and personal development.
- Contribuer au développement et à l'optimisation des modèles d'IA adaptés au matériel spécialisé de Graphcore.
- Travailler sur des systèmes à grande échelle où la performance est critique, en collaboration avec les équipes de développement logiciel et de recherche.
- Participer à la conception, à l'exécution et à la communication des résultats des expériences en apprentissage automatique.
- We seek engineers with strong technical skills and an understanding of AI model implementation at scale, eager to make a tangible impact in this rapidly evolving field
- If you're excited about advancing the next generation of AI models on cutting-edge hardware, we’d love to hear from you!
- Strong Python or C++ software development skills
- Enjoy cross-functional work collaborating with other teams
- Strong communicator - able to explain complex technical concepts to different audiences
- Experience in distributed training or inference of ML models across 64+ accelerators
- Developed deep understanding of performance bottlenecks and how to overcome them
- Bachelor/Master's/PhD or equivalent experience in Machine Learning, Computer Science, Maths, Data Science, or related field
- Capable of designing, executing and reporting from ML experiments
- Expertise in deep learning from model training to optimisation and evaluation
- Proficiency in deep learning frameworks like PyTorch/JAX
- Ability to move quickly in a dynamic environment
- Knowledge of cloud computing platforms
- Keen to present, publish and deliver talks in the AI community
- Have contributed to open-source projects or published research papers in relevant fields
- Experience writing C++/Triton/CUDA kernels for performance optimisation of ML models
- Familiarity with HPC systems and networking including Infiniband, NVLink, RoCE technologies
- Efficient computing based on low-precision arithmetic
- MLOps for Kubernetes-based clusters
- Building production systems with large language models
- Experience in one or more of:
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring