← back to jobs
> job detail
L
⚙️Data Engineer

#130529

Lifted, an Upwork Company™ · Bogotá, Bogota, Colombia
// classified as
Data Engineer (Pipelines, infra, ingestion, ETL.)
posted
1d ago
location
Bogotá, Bogota, Colombia
languages
java, python
tools
kafka, spark
> stack
javapythonkafkasparkairflow
> description

Job Description

Key Responsibilities

  • Design and develop scalable data processing solutions using Spark and Amazon EMR or comparable cloud-based data processing platforms.
  • Build and maintain batch and distributed data pipelines.
  • Develop software components for data transformation, feature preparation, and AI or machine learning workflow integration.
  • Collaborate with engineering, AI, and product teams to operationalize data-driven and model-enabled use cases.
  • Optimize data pipeline performance, cost efficiency, scalability, and production reliability.
  • Troubleshoot data and application issues across development and production environments.
  • Contribute to architecture discussions, technical documentation, and engineering standards.
  • Ensure solutions align with data quality, governance, and security expectations.

Qualifications

Must-Have Skills

  • 4+ years of software engineering or data engineering experience.
  • Strong experience with Spark and distributed data processing.
  • Experience with Amazon EMR or similar cloud-based data processing platforms.
  • Proficiency in Java, Python, or a related programming language.
  • Exposure to AI or machine learning workflows, model integration, or data preparation for intelligent systems.
  • Strong understanding of scalable data architecture and performance optimization.
  • Strong debugging and collaboration skills.
  • Comfortable delivering in evolving, data-intensive environments.
  • Ability to bridge software engineering and data engineering responsibilities.
  • Strong execution focus with practical architecture judgment.

Nice-to-Have Skills

  • Experience with Kafka, Airflow, data lakes, or data warehouse ecosystems.
  • Familiarity with MLOps, feature stores, or AI platform integration.
  • Experience with AWS-native services and observability tooling.
  • Enterprise experience strongly preferred.

Additional Information

Required Tools & Platforms

  • Apache Spark.
  • Amazon EMR or a comparable cloud-based distributed data processing platform.
  • Java, Python, or a related programming language.

Location, Time & Engagement

  • Remote contract role.
  • Candidates must be located in LATAM, excluding Mexico.
  • U.S. Central Time coverage is required.
  • Full-time allocation of approximately 40 hours per week.
  • Current contract end date is March 31, 2027.

Company Description

We are seeking an experienced Software/Data Engineer to design and deliver scalable data processing systems and AI-enabled workflows. This contract role sits at the intersection of software engineering and data engineering, with a strong focus on Spark, cloud-based distributed processing, production reliability, and data preparation for analytics and machine learning use cases