Senior Data Engineer, AI/ML Platform
Actively Hiring
Full-time Posted 16 days ago
Responsibilities
- check_circle Own and evolve our data ingestion pipelines and our web scraping engine as production systems: architecture, reliability, scalability, and performance, not just individual scripts.
- check_circle Design and build AI agent based extraction workflows (for example browser automation agents that navigate outlet websites, menus, and listing platforms) where they outperform traditional scraping and parsing approaches.
- check_circle Build evaluation frameworks and quality checks so AI-driven extraction and enrichment can be trusted at the same bar as our existing rule-based systems.
- check_circle Utilize AWS & Azure services effectively across both the traditional pipeline (Airflow, Spark, Glue) and the AI-driven components (model hosting, agent orchestration).
- check_circle Enhance our automated testing, data validation, and monitoring capabilities, including for AI components, which introduce new failure modes and cost profiles compared to deterministic scraping and rule-based logic.
- check_circle Collaborate with business and data stakeholders, including our data science team, to align requirements and expectations across the full pipeline.
- check_circle Drive engineering projects end to end and ensure successful outcomes.
- check_circle Drive continuous improvement in our engineering setup.
- check_circle Foster a collaborative and enjoyable team environment through code reviews, knowledge sharing, and daily interactions.
Basic qualifications
- Bachelor's and/or Master's degree in computer science or a related field.
- 5+ years of experience as a Data Engineer with strong software engineering skills, including ownership of production data pipelines and/or web scraping and crawling systems.
- Design, build, and operate scalable data pipelines on AWS that bring in data from external web sources and turn it into clean, queryable datasets with Playwright and/or BeautifulSoup.
- Experience with Airflow, Python, SQL, and Spark.
- Hands-on, practical experience integrating LLMs into production systems: prompt design as part of system design, evaluation, and cost/latency tradeoffs, not just experimenting in a notebook.
- Experience with AI agent frameworks (for example browser automation agents) or a strong interest in and aptitude for learning them.
- Manage infrastructure as code with Terraform on AWS (ECS, EMR, Glue, S3, SQS/SNS, IAM)
- Develop and maintain Spark jobs on EMR for batch ETL and enrichment at scale.
- Some exposure to classical ML/NLP concepts (classification, embeddings, matching) is a plus, since our data science team's pipeline uses these heavily.
- Result-driven and hands-on attitude.
- Proactive and courage to speak up.
- Ability to communicate in English, verbally and in writing.
- A competitive salary & pension plan.
- Enjoy 25 days of paid holiday per year and 8% holiday allowance.
- NS-Business card for an easy commute to our awesome NDSM-Werf office.
- A working-from-home allowance of €250 for a top-notch home office.
- A company laptop for enhanced productivity.
- Flexible work hours for a balanced lifestyle, with 2 days a week at our Amsterdam office.
- Possibly some great fun at our Friday drinks and team activities.
About the company
- check_circle Roamler Retail provides in-store and out-of-home insights and execution for FMCG brands and retailers across Europe, enabling clients to increase sales volumes while optimising operational efficiency.
- check_circle Roamler Tech delivers installation, maintenance, and repair services for large B2C service providers, improving customer experience while increasing the efficiency and flexibility of field execution.
Tags & Focus Areas
Fulltime Mlops Data Engineer Ai
About Roamler
Ready to Join the Team?
Apply once with DevFound — we route your profile to Roamler and keep you posted on matching AI roles.