#130529 - Software/Data Engineer - Spark, AWS EMR & AI

Upwork

Bogotá ciudad

Presencial

COP 279.520.000 - 372.694.000

Jornada completa

hace 6 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

No envíes un currículum genérico: crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Descripción de la vacante

Upwork is seeking an experienced Software/Data Engineer to design and deliver scalable data processing systems and AI-enabled workflows. This contract role sits at the intersection of software engineering and data engineering, with a strong focus on Spark, cloud-based distributed processing, production reliability, and data preparation for analytics and machine learning use cases.

You will design and develop scalable data processing solutions, build batch and distributed pipelines, and

Formación

  • 4+ years of software engineering or data engineering experience.
  • Strong experience with Spark and distributed data processing.
  • Experience with Amazon EMR or similar cloud-based data processing platforms.
  • Proficiency in Java, Python, or a related programming language.
  • Exposure to AI or machine learning workflows, model integration, or data preparation for intelligent systems.
  • Strong understanding of scalable data architecture and performance optimization.
  • Strong debugging and collaboration skills.
  • Comfortable delivering in evolving, data-intensive environments.
  • Ability to bridge software engineering and data engineering responsibilities.
  • Strong execution focus with practical architecture judgment.

Responsabilidades

  • Design and develop scalable data processing solutions using Spark and Amazon EMR or comparable cloud-based data processing platforms.
  • Build and maintain batch and distributed data pipelines.
  • Develop software components for data transformation, feature preparation, and AI or machine learning workflow integration.
  • Collaborate with engineering, AI, and product teams to operationalize data-driven and model-enabled use cases.
  • Optimize data pipeline performance, cost efficiency, scalability, and production reliability.
  • Troubleshoot data and application issues across development and production environments.
  • Contribute to architecture discussions, technical documentation, and engineering standards.
  • Ensure solutions align with data quality, governance, and security expectations.

Conocimientos

Spark
Distributed processing
Amazon EMR
Java
Python
AI workflows
Data architecture
Debugging
Collaboration
Data security

Herramientas

Airflow
Kafka
AWS

Descripción del empleo

We are seeking an experienced Software/Data Engineer to design and deliver scalable data processing systems and AI-enabled workflows. This contract role sits at the intersection of software engineering and data engineering, with a strong focus on Spark, cloud-based distributed processing, production reliability, and data preparation for analytics and machine learning use cases

Key Responsibilities
  • Design and develop scalable data processing solutions using Spark and Amazon EMR or comparable cloud-based data processing platforms.
  • Build and maintain batch and distributed data pipelines.
  • Develop software components for data transformation, feature preparation, and AI or machine learning workflow integration.
  • Collaborate with engineering, AI, and product teams to operationalize data-driven and model-enabled use cases.
  • Optimize data pipeline performance, cost efficiency, scalability, and production reliability.
  • Troubleshoot data and application issues across development and production environments.
  • Contribute to architecture discussions, technical documentation, and engineering standards.
  • Ensure solutions align with data quality, governance, and security expectations.
Must-Have Skills
  • 4+ years of software engineering or data engineering experience.
  • Strong experience with Spark and distributed data processing.
  • Experience with Amazon EMR or similar cloud-based data processing platforms.
  • Proficiency in Java, Python, or a related programming language.
  • Exposure to AI or machine learning workflows, model integration, or data preparation for intelligent systems.
  • Strong understanding of scalable data architecture and performance optimization.
  • Strong debugging and collaboration skills.
  • Comfortable delivering in evolving, data-intensive environments.
  • Ability to bridge software engineering and data engineering responsibilities.
  • Strong execution focus with practical architecture judgment.
Nice-to-Have Skills
  • Experience with Kafka, Airflow, data lakes, or data warehouse ecosystems.
  • Familiarity with MLOps, feature stores, or AI platform integration.
  • Experience with AWS-native services and observability tooling.
  • Enterprise experience strongly preferred.
Required Tools & Platforms
  • Apache Spark.
  • Amazon EMR or a comparable cloud-based distributed data processing platform.
  • Java, Python, or a related programming language.
Location, Time & Engagement
  • Remote contract role.
  • Candidates must be located in LATAM, excluding Mexico.
  • U.S. Central Time coverage is required.
  • Full-time allocation of approximately 40 hours per week.
  • Current contract end date is March 31, 2027.
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Remote Data Engineer - Spark, AWS EMR & AI Pipelines
Remote Data Engineer - Spark, AWS EMR & AI Pipelines

Upwork • Bogotá ciudad

Presencial
COP 279.520.000 - 372.694.000
Senior Data Engineer - Spark, EMR & AI Pipelines (Remote)
Senior Data Engineer - Spark, EMR & AI Pipelines (Remote)

AgileEngine • Colombia

Presencial
COP 372.694.000 - 559.041.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Senior Data Engineer: Spark, EMR & AI Workflows (Remote)
Senior Data Engineer: Spark, EMR & AI Workflows (Remote)

AgileEngine, LLC. • Pereira

Presencial
COP 120.000.000 - 180.000.000
Growth without limits
Competitive compensation
Flexibility: remote work
Senior Data Engineer: Spark & AI Pipelines (Remote)
Senior Data Engineer: Spark & AI Pipelines (Remote)

AgileEngine, LLC. • Sur

Presencial
COP 378.310.000 - 504.414.000
Growth budget
Competitive pay
Remote work
+3
Senior Data Engineer | Spark, EMR, AI Workflows (Remote)
Senior Data Engineer | Spark, EMR, AI Workflows (Remote)

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Presencial
COP 283.733.000 - 378.310.000
Growth without limits
Competitive compensation
100% remote with flexible hours
+2
Senior Software Engineer - Spark Pipelines (Remote)
Senior Software Engineer - Spark Pipelines (Remote)

AgileEngine • Perímetro Urbano Barranquilla

Presencial
COP 80.000.000 - 120.000.000
Growth opportunities
Competitive pay
Remote work / flexible hours
+3
Senior Software Engineer - Spark Data Pipelines + AI Remote
Senior Software Engineer - Spark Data Pipelines + AI Remote

INGEPSY • Sucre

Presencial
COP 100.440.000 - 200.880.000
Growth without limits
Competitive compensation
100% remote work
+3
Senior Data Engineer - Spark/EMR, Remote & AI Pipelines
Senior Data Engineer - Spark/EMR, Remote & AI Pipelines

AgileEngine • Sur

Presencial
COP 282.672.000 - 439.712.000
Growth without limits
Competitive compensation
Remote work 100%
+3
Senior Data Engineer: Spark, EMR & AI Workflows (Remote)
Senior Data Engineer: Spark, EMR & AI Workflows (Remote)

AgileEngine, LLC. • Metropolitana

Presencial
COP 283.733.000 - 378.310.000
Growth without limits
Competitive compensation
Flexibility — 100% remote
+3
Remote Data Engineer: Spark, EMR & AI
Remote Data Engineer: Spark, EMR & AI

Lifted, an Upwork Company™ • Bogotá

Presencial
COP 221.758.000 - 348.476.000