Databricks Engineer

RevStar Consulting Inc

Norte

A distancia

COP 434.958.000 - 590.300.000

Jornada completa

14 días+
Generador de candidaturas

Transforma esta oferta en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Remote-first environment
Health coverage
401(k) retirement plan

Descripción de la vacante

RevStar Consulting Inc is seeking a Databricks Engineer to build, optimize, and deploy scalable data pipelines and MLOps frameworks for enterprise clients across AWS, Azure, and GCP. You will work with data architects, scientists, and client leaders to turn complex data into production-ready AI models in a remote setting.

The role centers on Lakehouse architecture, CI/CD, IaC, and data governance for cloud-agnostic solutions, with a focus on performance, monitoring, and security.

Formación

  • 3+ years of hands-on data engineering experience with big data and cloud-native architectures.
  • 2+ years with Databricks (Spark, Delta Lake, MLflow).
  • Databricks Certifications mandatory: Databricks Certified Data Engineer Associate or higher.
  • Proficiency in Python, SQL, and Spark-based frameworks.
  • Experience in developing and optimizing large-scale ETL/ELT pipelines.
  • Strong understanding of Lakehouse architecture and cloud-agnostic data solutions.
  • Familiarity with CI/CD pipelines and Infrastructure-as-Code (IaC) for Databricks (Terraform, Databricks CLI).
  • Knowledge of data governance, security, and compliance best practices.
  • Experience working in Agile development environments, DevOps/MLOps.

Responsabilidades

  • Design, build, and optimize scalable ETL/ELT pipelines across multi-cloud environments (AWS S3, Azure Data Lake, GCS).
  • Tune Spark jobs for low latency and high throughput with cost efficiency.
  • Implement CI/CD pipelines and IaC for automated Databricks deployments.
  • Build automated monitoring, alerting, and data quality validation frameworks.
  • Collaborate with ML engineers to productionize feature pipelines and MLflow tracking.
  • Operationalize AI/ML models in secure, scalable production environments.

Conocimientos

Databricks
Python
SQL
Spark
Delta Lake
MLflow
ETL/ELT pipelines
CI/CD
IaC
DevOps/MLOps

Herramientas

Terraform
Databricks CLI
AWS
Azure
GCP

Descripción del empleo

  • Reports To: Data & AI Practice Lead
  • Location: Remote (US-Based / Eastern or Central Time Zone Preferred)
  • Employment Type: Contract

Ready to build greenfield Lakehouse solutions at the bleeding edge of AI and big data? RevStar is an innovation shop and official Databricks Partner launching a dedicated, cloud-agnostic Data, ML, and AI practice. We are seeking a high-caliber Databricks Engineer to build, optimize, and deploy high-performance data pipelines and MLOps frameworks for enterprise clients across AWS, Azure, and GCP. In this role, you will work directly with data architects, scientists, and client leaders to turn complex data into scalable, production-ready AI models. Above all, the ideal candidate embodies RevStar’s core values:

Self-Mastery: We hold a high bar for how we think, communicate, and improve.

Ownership: We own outcomes, not just effort.

Shared Destiny: We rise or fall together.

Your Impact Pillars

Your technical contributions are organized into four strategic pillars:

1. Lakehouse Architecture & Pipeline Engineering

  • Design, build, and optimize scalable ETL/ELT pipelines using Apache Spark and Delta Lake across multi-cloud environments (AWS S3, Azure Data Lake, GCS).
  • Implement robust Lakehouse architectures that seamlessly process both structured and unstructured data at enterprise scale.
  • Automate data ingestion and storage workflows to support downstream analytics and real-time operational reporting.

2. Performance Optimization & Automation

  • Fine-tune Spark jobs for low latency, high throughput, and maximum cloud cost-efficiency.
  • Implement CI/CD pipelines and Infrastructure-as-Code (Terraform, Databricks CLI) for automated deployments.
  • Build automated monitoring, alerting, and data quality validation frameworks to guarantee pipeline reliability.

3. MLOps & AI Integration

  • Partner with ML engineers and data scientists to build production-grade feature engineering pipelines.
  • Support model training, tracking, versioning, and deployment inside Databricks using MLflow.
  • Operationalize AI/ML models into secure, scalable production environments for client applications.

4. Data Governance & Client Excellence

  • Enforce enterprise data security, access controls, and compliance standards (GDPR, HIPAA, SOC 2).
  • Establish best practices for data lineage, metadata management, and operational documentation.
  • Collaborate with client-facing stakeholders to align technical implementations with critical business outcomes.

Must-Have:

  • 3+ years of hands-on experience in data engineering, with a focus on big data processing and cloud-native architectures.
  • 2+ years of hands-on experience with Databricks, including Apache Spark, Delta Lake, and MLflow.
  • Databricks Certifications (Mandatory):
    • Databricks Certified Data Engineer Associate (or higher)
  • Proficiency in Python, SQL, and Spark-based frameworks.
  • Experience in developing and optimizing large-scale ETL/ELT pipelines.
  • Strong understanding of Lakehouse architecture and cloud-agnostic data solutions.
  • Familiarity with CI/CD pipelines and Infrastructure-as-Code (IaC) for Databricks (e.g., Terraform, Databricks CLI).
  • Knowledge of data governance, security, and compliance best practices.
  • Experience working in Agile development environments, following DevOps/MLOps best practices.

Nice-to-Have:

  • Additional Databricks Certifications (e.g., Databricks Certified Machine Learning Associate).
  • Experience with real-time streaming solutions (e.g., Kafka, Kinesis, Event Hub).
  • Familiarity with cloud storage and orchestration tools (e.g., Apache Airflow, Prefect).
  • Background in AI/ML integration within Databricks, assisting in feature engineering and model deployment.
  • Experience working in client-facing roles or consulting environments.

Benefits for Full-Time W2 Positions:

  • Paid Time Off – Take the time you need to recharge and stay productive.
  • Remote-First Working Environment – Collaborate from anywhere while staying connected with our global team.
  • Comprehensive Health Coverage – Medical, Dental, Vision
  • 401(k) Retirement Plan – Plan for your future with access to a company-sponsored 401(k) program.
  • Annual Learning & Development Stipend – Invest in your skills with conferences, certifications, or courses.
  • Peer Mentorship & Coaching – Learn from experienced engineers, product managers, and architects to accelerate your growth.
  • Professional Growth Opportunities – Exposure to cutting-edge AWS GenAI, data, and cloud technologies across diverse industries.
  • Company Outings & Volunteer Opportunities – Build relationships and give back to the community.
  • Collaborative, Innovative Culture – Work alongside top talent in a fast-paced, supportive environment that values curiosity and initiative.

Equal Opportunity Employment

At RevStar, we don’t just accept differences — we celebrate them, we support them, and we thrive on them for the benefit of our employees, our customers, and our community. RevStar is proud to be an equal opportunity workplace.

Reasonable Accommodations

RevStar is committed to providing access, equal opportunity and reasonable accommodation for individuals with disabilities in employment, its services, programs, and activities. To request reasonable accommodation, contact HR at hr@revstarconsulting.com.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Databricks Engineer - Lakehouse & MLOps (Remote, Contract)
Databricks Engineer - Lakehouse & MLOps (Remote, Contract)

RevStar Consulting Inc • Norte

A distancia
COP 434.958.000 - 590.300.000
Remote-first environment
Health coverage
401(k) retirement plan
1104 | Senior Data (Databricks) Engineer
1104 | Senior Data (Databricks) Engineer

Intetics • Colombia

Presencial
COP 108.000.000 - 180.000.000
Data Platform Engineer ID90126
Data Platform Engineer ID90126

AgileEngine, LLC. • Cartagena de Indias

Presencial
COP 40.000.000 - 75.000.000
Professional growth
Competitive USD-based compensation
Databricks Data Engineer: Lakehouse Pipelines & PySpark
Databricks Data Engineer: Lakehouse Pipelines & PySpark

Perficient • Colombia

Presencial
Confidential
Data Engineer ID89384
Data Engineer ID89384

AgileEngine, LLC. • Bogotá ciudad

Híbrido
COP 404.013.000 - 606.020.000
Professional growth
Competitive compensation
Exciting projects
+1
Data Engineer ID89384
Data Engineer ID89384

AgileEngine, LLC. • Cartagena de Indias

Presencial
COP 202.007.000 - 303.010.000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Data Engineer ID89384
Data Engineer ID89384

AgileEngine, LLC. • Capital

Híbrido
COP 404.013.000 - 538.684.000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Databricks Architect | Latam
Databricks Architect | Latam

Cuesta Partners, LLC • Perímetro Urbano Barranquilla, Bogotá ciudad, Sur, Medellín

Híbrido
COP 180.000.000 - 300.000.000
Health plan
Hybrid work
Performance bonus
+3
Data Engineer ID89384
Data Engineer ID89384

AgileEngine, LLC. • Medellín

Híbrido
COP 12.000.000 - 18.000.000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Senior Data Lead Engineer
Senior Data Lead Engineer

UST España & Latam • Colombia

Presencial
COP 256.166.288 - 329.356.656