← Back to job listings
RO
Senior Data Engineer (AI/ML Platform)
Roamler · Amsterdam, Netherlands
About The Role
Join our AI/ML platform team as a Senior Data Engineer. You will take ownership of the pipelines and systems behind our outlet-data product suite, working closely with data scientists and driving engineering projects end to end. You will design and build AI agent-based extraction workflows, enhance automated testing and monitoring capabilities, and collaborate with business and data stakeholders. Enjoy benefits such as 33 days of holiday, a company pension scheme, travel allowance, life insurance, medical insurance, and a bonus.
- Ownership and evolution of data ingestion pipelines and web scraping engine as production systems, focusing on architecture, reliability, scalability, and performance.
- Designing and building AI agent-based extraction workflows that outperform traditional scraping and parsing approaches, and building evaluation frameworks and quality checks.
- Collaboration with business and data stakeholders to align requirements and expectations across the full pipeline, driving engineering projects end to end.
- Manage infrastructure as code with Terraform on AWS (ECS, EMR, Glue, S3, SQS/SNS, IAM)
- Bachelor's and/or Master's degree in computer science or a related field
- Proactive and courage to speak up
- If you are not quite sure or feel that you don't meet 100% of the requirements, but are excited about us, apply anyway! Sometimes, the best combinations are a bit more unexpected
- Hands-on, practical experience integrating LLMs into production systems: prompt design as part of system design, evaluation, and cost/latency tradeoffs, not just experimenting in a notebook
- Ability to communicate in English, verbally and in writing
- Experience with AI agent frameworks (for example browser automation agents) or a strong interest in and aptitude for learning them
- Result-driven and hands-on attitude
- Experience with Airflow, Python, SQL, and Spark
- Develop and maintain Spark jobs on EMR for batch ETL and enrichment at scale
- 5+ years of experience as a Data Engineer with strong software engineering skills, including ownership of production data pipelines and/or web scraping and crawling systems
- Design, build, and operate scalable data pipelines on AWS that bring in data from external web sources and turn it into clean, queryable datasets with Playwright and/or BeautifulSoup
- Some exposure to classical ML/NLP concepts (classification, embeddings, matching) is a plus, since our data science team's pipeline uses these heavily
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring