Back to jobs

Embedded Data Engineer - ML

London, Greater London, United Kingdom

Pay
Salary not listed in the saved posting
Work setup
Unconfirmed
Employment
Unconfirmed
Apply at Trainline PLC

Tools in this posting

  • Python
  • AWS
  • dbt
  • Docker
  • Spark
  • Terraform
  • Airflow
  • SQL
  • Iceberg
Source — Tool mentions in context
We'd love to hear from you if you have...🔍 - Working knowledge of Python and SQL. - Experience building data pipelines for downstream machine learning workloads, including feature engineering and model training workflows.
- Design and build scalable data pipelines, data models, and feature stores that support analytics and machine learning workloads within the ML domain. - Deploy and maintain cloud-native data applications on AWS, using CI/CD pipelines to automate builds, testing, and releases. - Maintain the technical quality, performance, and reliability of production data pipelines through strong observability and engineering best practices.
- Comfort with data modelling and building efficient data marts and warehouses in the cloud. - Experience building data pipelines using tools such as Spark and Airflow, or similar technologies, within a cloud environment such as AWS. - Familiarity with both real-time and batch data workloads, along with modern data transformation and orchestration patterns.
Introducing the Embedded Data Engineering in ML Team 👋 At the heart of our Data and ML teams, embedded Data Engineers create the pipelines and tables that power business critical dashboards, enable self-service analytics, and fuel advanced machine learning models and real-time data products. Working with tools like DBT, Spark, and Airflow, you'll transform high volume raw event data into user-friendly, high impact datasets that support machine learning use cases across the business. As an Embedded Data Engineer in ML, you'll sit within the Machine Learning team, working day to day with Machine Learning Engineers and Data Scientists to build reliable datasets for ML use cases. You'll also have access to Trainline's wider Data Engineering, Data Platform, and analytics community, working alongside other embedded Data Engineers in ML, including senior and principal engineers.
- Ideally, you may also have experience with parallel or distributed training frameworks such as Ray, or with modern data formats such as Parquet and Iceberg. - It would also be helpful if you have some experience with Infrastructure as Code (Terraform) and containerisation (Docker) to support automated, standardised deployments. - You may also have contributed to or maintained CI/CD pipelines (such as Jenkins or GitHub Actions) as part of production grade data systems, and enjoy solving complex data problems collaboratively.
- Familiarity with both real-time and batch data workloads, along with modern data transformation and orchestration patterns. - Ideally, you may also have experience with parallel or distributed training frameworks such as Ray, or with modern data formats such as Parquet and Iceberg. - It would also be helpful if you have some experience with Infrastructure as Code (Terraform) and containerisation (Docker) to support automated, standardised deployments.

About Trainline PLC

At Trainline, our purpose is to empower greener travel choices, connecting people and places.

In the employer’s words · Read in context

Job description

View original posting ↗

About us

At Trainline, our purpose is to empower greener travel choices, connecting people and places. Trainline enables millions of travellers to find and book the best value tickets across carriers, fares, and journey options through our highly rated mobile app, website, and B2B partner channels. 

Great journeys start with Trainline 🚄 

We’re Europe’s leading independent rail platform, helping millions of travellers find and book the best-value rail and coach journeys across our app, website and partner channels.

Our job is to make the green travel choice the best choice. By building a better train travel experience, we help more people choose rail - creating a positive impact for customers, our business and the planet.

We’re a team of more than 1,000 Trainliners from over 50 nationalities, working across London, Paris, Barcelona, Milan, Edinburgh and Madrid. Now is a brilliant time to join us and help shape the future of travel.

Introducing the Embedded Data Engineering in ML Team 👋

At the heart of our Data and ML teams, embedded Data Engineers create the pipelines and tables that power business critical dashboards, enable self-service analytics, and fuel advanced machine learning models and real-time data products. Working with tools like DBT, Spark, and Airflow, you'll transform high volume raw event data into user-friendly, high impact datasets that support machine learning use cases across the business.

As an Embedded Data Engineer in ML, you'll sit within the Machine Learning team, working day to day with Machine Learning Engineers and Data Scientists to build reliable datasets for ML use cases. You'll also have access to Trainline's wider Data Engineering, Data Platform, and analytics community, working alongside other embedded Data Engineers in ML, including senior and principal engineers.

In this role as the Embedded Data Engineer (ML), you will...🚄

  • Design and build scalable data pipelines, data models, and feature stores that support analytics and machine learning workloads within the ML domain.

  • Deploy and maintain cloud-native data applications on AWS, using CI/CD pipelines to automate builds, testing, and releases.

  • Maintain the technical quality, performance, and reliability of production data pipelines through strong observability and engineering best practices.

  • Collaborate closely with Machine Learning Engineers and Data Scientists to build reliable, well-structured datasets that power ML use cases.

  • Work with the wider Data Engineering, Data Platform, and analytics community to share knowledge and align on best practices across teams.

We'd love to hear from you if you have...🔍

  • Working knowledge of Python and SQL.

  • Experience building data pipelines for downstream machine learning workloads, including feature engineering and model training workflows.

  • Comfort with data modelling and building efficient data marts and warehouses in the cloud.

  • Experience building data pipelines using tools such as Spark and Airflow, or similar technologies, within a cloud environment such as AWS.

  • Familiarity with both real-time and batch data workloads, along with modern data transformation and orchestration patterns.

  • Ideally, you may also have experience with parallel or distributed training frameworks such as Ray, or with modern data formats such as Parquet and Iceberg.

  • It would also be helpful if you have some experience with Infrastructure as Code (Terraform) and containerisation (Docker) to support automated, standardised deployments.

  • You may also have contributed to or maintained CI/CD pipelines (such as Jenkins or GitHub Actions) as part of production grade data systems, and enjoy solving complex data problems collaboratively.

More information:

Enjoy fantastic perks like private healthcare & dental insurance, a generous work from abroad policy, 2-for-1 share purchase plans, an EV Scheme to further reduce carbon emissions, extra festive time off, and excellent family-friendly benefits. 

We prioritise career growth with clear career paths, transparent pay bands, personal learning budgets, and regular learning days. Jump on board and supercharge your career from day one! 

We're operating a hybrid model and ask that Trainliners work from the office a minimum of 60% of their time over a 12-week period. We also have a 28-day Work from Abroad policy.

Our values represent the things that matter most to us and what we live and breathe everyday, in everything we do: 

  • 💭 Think Big - We're building the future of rail 

  • ✔️ Own It - We focus on every customer, partner and journey 

  • 🤝  Travel Together - We're one team 

  • ♻️ Do Good - We make a positive impact 

We know that having a diverse team makes us better and helps us succeed. And we mean all forms of diversity - gender, ethnicity, sexuality, disability, nationality and diversity of thought. That's why we're committed to creating inclusive places to work, where everyone belongs and differences are valued and celebrated.

Interested in finding out more about what it's like to work at Trainline? Why not check us out on LinkedIn, Instagram and Glassdoor! 

Your next step

  • Have your CV and examples of relevant work ready.
  • Check the listed location, eligibility and core experience before starting.
  • Ask the employer about the salary range before committing time to the process.

Complete your application on jobs.ashbyhq.com. The employer’s form will show what is required.

Already applied? Track this application

Source & posting history

View original posting ↗

Source notes

Source excerpts

Selected passages from the saved posting. Check the full description for conditions and exceptions.

Pay

No pay amount identified in the saved description.

Location & working pattern

London, Greater London, United Kingdom

We prioritise career growth with clear career paths, transparent pay bands, personal learning budgets, and regular learning days. Jump on board and supercharge your career from day one! We're operating a hybrid model and ask that Trainliners work from the office a minimum of 60% of their time over a 12-week period. We also have a 28-day Work from Abroad policy. Our values represent the things that matter most to us and what we live and breathe everyday, in everything we do:
Work authorization

No clear work-authorization passage found. Eligibility is unconfirmed.

Status in our records
Active
First seen by us
Sep 10, 2026
Recorded sightings
20
Last seen by us
Oct 9, 2026

These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.

Report an error

See how this role fits your experience

Add your resume to compare the role’s scope, tools and requirements with your experience.

Find answers in the posting

AI
How answers work

AI selects complete passages from this posting. Check them for conditions and exceptions.

Uses this posting and your question. No profile needed.