Machine Learning Research Engineer (LLMs & AI Systems)
Tenstorrent · Toronto, Canada
About The Role
Join Tenstorrent as a Machine Learning Research Engineer, where you will lead research and development efforts focused on optimizing large language models (LLMs) and AI systems. You will train, evaluate, and optimize state-of-the-art AI models on Tenstorrent hardware, improve performance through various techniques, investigate system bottlenecks, and collaborate cross-functionally to drive performance improvements. This role offers the opportunity to learn about optimizing AI models on custom AI accelerators, deploying and scaling large-scale ML systems, and the collaboration between hardware, compiler, kernel, and ML teams.
- Lead research and development efforts focused on LLM training and inference optimization.
- Train, evaluate, and optimize state-of-the-art AI models on Tenstorrent hardware.
- Investigate system bottlenecks and collaborate cross-functionally to drive performance improvements.
- Strong Python and PyTorch experience developing and training deep learning models
- 4+ years of industry and/or academic experience in ML research and LLM development
- PhD, published research, or experience with speculative decoding is highly valued
- Hands-on experience training large-scale machine learning models
- Deep understanding of ML architectures, LLM training, and inference optimization
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring