← Back to job listings
XA
Member of Technical Staff (RL Inference)
xAI · Palo Alto, United States
About The Role
Join the RL infrastructure team at xAI as a Member of Technical Staff. In this role, you will design and optimize our inference stack for various RL workloads, analyze and address performance bottlenecks, and collaborate with the modeling team to implement novel RL techniques. Enjoy comprehensive health insurance, flexible vacation, visa sponsorship, and a 401(k) plan.
- Conception et optimisation de l'infrastructure d'inférence pour les charges de travail RL.
- Analyse, profilage et résolution des goulets d'étranglement de performance dans les systèmes RL à grande échelle.
- Collaboration avec l'équipe de modélisation pour mettre en œuvre efficacement de nouvelles techniques et algorithmes RL.
- All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates
- Proficiency in programming languages such as Python, C++ and/or Rust; frameworks such as PyTorch, Jax, CUDA
- Experience in LLM inference
- Willingness to dive deep and solve hardcore problems at all levels of the stack
- Experience in building, debugging, and optimizing efficiency of large-scale distributed systems
- Strong knowledge in quantization and numerics in LLM inference and training
- Experience in developing inference engines, e.g. SGLang, vLLM
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring