AI Data Engineer
Prague
- Pay
- Salary not listed in the saved posting
- Work setup
- Unconfirmed
- Employment
- Unconfirmed
What you’ll work on
Full postingDesign, develop, and maintain scalable data ingestion, transformation, and enrichment pipelines supporting AI and analytics initiatives.
Lead initiatives to replace manual, siloed, or obsolete data workflows with automated, cloud-enabled, and data-driven solutions.
Develop data processing solutions using Python, Spark, SQL, and distributed computing frameworks.
From the employer’s posting
What You Will Be Doing Design, develop, and maintain scalable data ingestion, transformation, and enrichment pipelines supporting AI and analytics initiatives. Assess, modernize, and migrate legacy data platforms, tools, and operational processes to scalable, governed, and AI-ready data infrastructures.
Assess, modernize, and migrate legacy data platforms, tools, and operational processes to scalable, governed, and AI-ready data infrastructures. Lead initiatives to replace manual, siloed, or obsolete data workflows with automated, cloud-enabled, and data-driven solutions. Design and execute migration strategies for data assets stored in legacy applications, databases, file systems, and on-premise environments while ensuring data quality and business continuity.
Refactor and optimize existing ETL/ELT pipelines, data models, and integration processes to align with modern architectural standards and AI platform requirements. Develop data processing solutions using Python, Spark, SQL, and distributed computing frameworks. Design and implement data architectures supporting LLM-based applications, Retrieval-Augmented Generation (RAG), semantic search, and knowledge management platforms.
What you’ll bring
All qualificationsCore experience
- Bachelor's or Master's degree in Computer Science, Data Science, or a related field.
- Strong proficiency in programming languages such as Scala, Python, or Java.
- Familiarity with APIs, data integration frameworks, and event-driven architectures.
- Knowledge of data security, encryption, access controls, and compliance requirements.
- Understanding of modern data architectures, including data lakes, lakehouses, and data mesh concepts.
- Familiarity with AI and ML workflows, including training data preparation, feature engineering, and model-serving data pipelines
Qualification wording
Bachelor's or Master's degree in Computer Science, Data Science, or a related field.
Strong proficiency in programming languages such as Scala, Python, or Java.
Familiarity with APIs, data integration frameworks, and event-driven architectures.
Knowledge of data security, encryption, access controls, and compliance requirements.
Understanding of modern data architectures, including data lakes, lakehouses, and data mesh concepts.
Familiarity with AI and ML workflows, including training data preparation, feature engineering, and model-serving data pipelines
Tools in this posting
- Python
- Scala
- SQL
- Spark
- Java
Source — Tool mentions in context
Refactor and optimize existing ETL/ELT pipelines, data models, and integration processes to align with modern architectural standards and AI platform requirements. Develop data processing solutions using Python, Spark, SQL, and distributed computing frameworks. Design and implement data architectures supporting LLM-based applications, Retrieval-Augmented Generation (RAG), semantic search, and knowledge management platforms.
Proven at least 5 years of working experience as a Big Data Engineer or similar. Strong proficiency in programming languages such as Scala, Python, or Java. Familiarity with APIs, data integration frameworks, and event-driven architectures.
At least 5 years of experience in migrating legacy data platforms, repositories, and operational workflows to modern data architectures while ensuring business continuity and data integrity. Building data solutions using Python, SQL, Spark, and distributed computing frameworks experience. Knowledge of data security, encryption, access controls, and compliance requirements.
Job description
Position Purpose:
We are seeking a passionate and motivated AI Data Engineer to join our international team as part of a strategic expansion of our AI and Data capabilities.
In this role, you will be responsible for designing, building, and maintaining the data foundations that power foundational applications, advanced analytics, machine learning, Generative AI, and Large Language Model (LLM) applications across the organization. You will work at the intersection of data engineering, AI platform development, and data governance, ensuring that high-quality, trusted, and scalable data is available to support AI-driven innovation.
Your work will play a critical role in enabling next-generation AI solutions by building robust data pipelines, optimizing large-scale data architectures, and creating AI-ready datasets that support model training, retrieval-augmented generation (RAG), semantic search, and intelligent automation initiatives.
As an AI Data Engineer, you will collaborate closely with Data Engineers, Machine Learning Engineers, Data Scientists, Data Architects, Product teams, and business stakeholders to ensure that data assets are reliable, accessible, secure, and aligned with both business and technical requirements.
What You Will Be Doing
Design, develop, and maintain scalable data ingestion, transformation, and enrichment pipelines supporting AI and analytics initiatives.
Assess, modernize, and migrate legacy data platforms, tools, and operational processes to scalable, governed, and AI-ready data infrastructures.
Lead initiatives to replace manual, siloed, or obsolete data workflows with automated, cloud-enabled, and data-driven solutions.
Design and execute migration strategies for data assets stored in legacy applications, databases, file systems, and on-premise environments while ensuring data quality and business continuity.
Refactor and optimize existing ETL/ELT pipelines, data models, and integration processes to align with modern architectural standards and AI platform requirements.
Develop data processing solutions using Python, Spark, SQL, and distributed computing frameworks.
Design and implement data architectures supporting LLM-based applications, Retrieval-Augmented Generation (RAG), semantic search, and knowledge management platforms.
Develop and manage metadata, lineage, and cataloging solutions to improve data discoverability and governance.
Collaborate with Machine Learning Engineers and Data Scientists to provide high-quality feature stores, training datasets, and inference data pipelines.
Support the creation and maintenance of Data-as-a-Service (DaaS) platforms and reusable data products.
Implement automated data quality checks, validation frameworks, and monitoring solutions to ensure data integrity and reliability.
Contribute to AI platform engineering initiatives, including model-serving infrastructure, feature management, and AI data lifecycle management.
Ensure compliance with data privacy regulations, security standards, and organizational governance policies.
Define and implement best practices to ensure migrated systems comply with enterprise standards for data governance, security, lineage, observability, and documentation.
Participate in the design of AI governance processes related to data sourcing, quality management, access control, and responsible AI.
Collaborate with business stakeholders to identify opportunities for process automation, technical debt reduction, and platform modernization.
Stay informed on emerging technologies in data engineering, Generative AI, knowledge graphs, vector search, and AI platform architectures.
Participate in agile ceremonies including sprint planning, solution design reviews, architecture discussions, and continuous improvement initiatives.
What You Need for this Position
Bachelor's or Master's degree in Computer Science, Data Science, or a related field.
Proven at least 5 years of working experience as a Big Data Engineer or similar.
Strong proficiency in programming languages such as Scala, Python, or Java.
Familiarity with APIs, data integration frameworks, and event-driven architectures.
At least 5 years of experience in migrating legacy data platforms, repositories, and operational workflows to modern data architectures while ensuring business continuity and data integrity.
Building data solutions using Python, SQL, Spark, and distributed computing frameworks experience.
Knowledge of data security, encryption, access controls, and compliance requirements.
Understanding of modern data architectures, including data lakes, lakehouses, and data mesh concepts.
Familiarity with AI and ML workflows, including training data preparation, feature engineering, and model-serving data pipelines
Experience with Git, CI/CD pipelines, and infrastructure-as-code practices.
Strong analytical and problem-solving capabilities, with the ability to navigate complex and ambiguous challenges.
Passion for AI, data innovation, and continuous improvement.
Ability to assess existing processes and identify opportunities for automation, modernization, and operational efficiency.
Excellent communication and stakeholder management skills.
Ability to prioritize tasks, manage multiple projects simultaneously, and deliver high-quality results within defined timelines.
Ability to collaborate effectively across Data Engineering, Data Science, AI, Product, Architecture, and Business teams.
Strong organizational skills and attention to detail.
Comfortable working in fast-paced, multicultural, and international environments.
Curiosity and the ability to rapidly learn new technologies, platforms, and engineering practices.
The ability to thrive in a fast-paced environment and take on new responsibilities quickly.
Your next step
- Have your CV and examples of relevant work ready.
- Check the listed location, eligibility and core experience before starting.
- Ask the employer about the salary range before committing time to the process.
Complete your application on solera.wd5.myworkdayjobs.com. The employer’s form will show what is required.
Already applied? Track this application
Source & posting history
Source notes
Source excerptsSelected passages from the saved posting. Check the full description for conditions and exceptions.
- Pay
No pay amount identified in the saved description.
- Location & working pattern
Prague
Working pattern and location restrictions need checking in the full posting.
- Work authorization
No clear work-authorization passage found. Eligibility is unconfirmed.
- Status in our records
- Active
- First seen by us
- Oct 2, 2026
- Recorded sightings
- 1
These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.
Report an errorSee how this role fits your experience
Add your resume to compare the role’s scope, tools and requirements with your experience.