Skip to content
← Back to job listings

Senior Data Engineer (AI/ML Platform)

Roamler · Amsterdam, Netherlands

External listingfull-time26 days ago

About The Role

Join our AI/ML platform team as a Senior Data Engineer. You will take ownership of the pipelines and systems behind our outlet-data product suite, working closely with data scientists and driving engineering projects end to end. You will design and build AI agent-based extraction workflows, enhance automated testing and monitoring capabilities, and collaborate with business and data stakeholders. Enjoy benefits such as 33 days of holiday, a company pension scheme, travel allowance, life insurance, medical insurance, and a bonus.

  • Ownership and evolution of data ingestion pipelines and web scraping engine as production systems, focusing on architecture, reliability, scalability, and performance.
  • Designing and building AI agent-based extraction workflows that outperform traditional scraping and parsing approaches, and building evaluation frameworks and quality checks.
  • Collaboration with business and data stakeholders to align requirements and expectations across the full pipeline, driving engineering projects end to end.
  • Manage infrastructure as code with Terraform on AWS (ECS, EMR, Glue, S3, SQS/SNS, IAM)
  • Bachelor's and/or Master's degree in computer science or a related field
  • Proactive and courage to speak up
  • If you are not quite sure or feel that you don't meet 100% of the requirements, but are excited about us, apply anyway! Sometimes, the best combinations are a bit more unexpected
  • Hands-on, practical experience integrating LLMs into production systems: prompt design as part of system design, evaluation, and cost/latency tradeoffs, not just experimenting in a notebook
  • Ability to communicate in English, verbally and in writing
  • Experience with AI agent frameworks (for example browser automation agents) or a strong interest in and aptitude for learning them
  • Result-driven and hands-on attitude
  • Experience with Airflow, Python, SQL, and Spark
  • Develop and maintain Spark jobs on EMR for batch ETL and enrichment at scale
  • 5+ years of experience as a Data Engineer with strong software engineering skills, including ownership of production data pipelines and/or web scraping and crawling systems
  • Design, build, and operate scalable data pipelines on AWS that bring in data from external web sources and turn it into clean, queryable datasets with Playwright and/or BeautifulSoup
  • Some exposure to classical ML/NLP concepts (classification, embeddings, matching) is a plus, since our data science team's pipeline uses these heavily

This is an external listing. JobSpring does not represent or verify the employer. Report this listing