← Back to job listings
WI
AI Applied Scientist
Wizard · United States
About The Role
Join our team as an AI Applied Scientist, where you will define and evolve accuracy metrics across the full shopping experience, design and run experiments to measure improvements and regressions, and build and maintain evaluation datasets, benchmarks, and scoring frameworks. You will work closely with ML Engineers to validate model changes and guide iteration, identify failure modes and edge cases, and drive improvements through data. This role offers opportunities for career growth in depth, breadth, and leadership tracks.
- Definir y evolucionar métricas de precisión a lo largo de toda la experiencia de compra, incluyendo recuperación, clasificación, recomendaciones y resultados.
- Diseñar y ejecutar experimentos para medir mejoras y regresiones, y construir y mantener conjuntos de datos de evaluación, benchmarks y marcos de puntuación.
- Mejorar los jueces de LLM que impulsan nuestra pipeline de evaluación, incluyendo el prompting, la calibración y el fine-tuning donde sea necesario.
- Proven ability to operate in ambiguity: defining problems, not just solving pre-defined ones
- 5+ years in Applied ML, AI Research, or Applied Science (PhD or equivalent depth strongly preferred)
- Hands-on experience evaluating modern AI/ML systems: LLMs, agents, ranking, or recommendations
- Clear, structured communication that influences across ML, engineering, and product
- Strong experimentation foundations: A/B testing, causal inference, statistical rigor
- Direct experience with LLM-based systems: judge models, RAG, prompt engineering, fine-tuning, RLHF, or similar
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring