Cloud Engineer

Power Staffing Solutions

Atlanta (GA)

On-site

USD 130,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology staffing firm in Atlanta is seeking a Senior Cloud Engineer specializing in Data Platforms. The role involves designing and managing cloud infrastructure for GenAI applications with a focus on AWS and Databricks. Ideal candidates will have extensive experience in cloud architecture and disaster recovery strategies, along with a strong grasp of Terraform and serverless solutions. This full-time position offers an opportunity to lead innovative cloud projects.

Qualifications

  • 5+ years of hands-on experience with AWS services.
  • 2+ years of hands-on experience with Databricks.
  • Strong proficiency with Terraform for infrastructure automation.

Responsibilities

  • Design and maintain cloud resources across AWS, Azure, and Google Cloud.
  • Implement disaster recovery strategies and manage high availability.
  • Build and maintain a self-service platform for GenAI applications.

Skills

AWS services
Databricks
Terraform
Containerization (Docker, Kubernetes)
Python
Event-driven architecture
Serverless frameworks

Education

Bachelor's degree in Computer Science or related field

Tools

AWS
Azure
Google Cloud Platform

Job description

Get AI-powered advice on this job and more exclusive features.

We are seeking a highly skilled Cloud Engineer (Senior, Data Platforms) to join our team and lead the design, implementation, and management of cloud infrastructure for innovative GenAI applications. This role will be instrumental in building a robust platform that enables rapid experimentation and deployment while maintaining enterprise‑grade security and reliability.

We are a growing data product company working primarily with Fortune 500 organizations. Our focus is on delivering digital solutions that accelerate business growth by creating value through innovation.

Job Responsibilities

  • Design, provision, and maintain cloud resources across AWS (primary), with the ability to work in Azure and Google Cloud environments.
  • Manage end-to-end infrastructure for full-stack GenAI applications including:
  • Security groups and IAM policies
  • VPC architecture and network design
  • Container orchestration (ECS, EKS, Lambda)
  • Storage solutions (S3, EFS)
  • CDN configuration (CloudFront)
  • DNS management (Route53)
  • Load balancing and auto-scaling

Data & AI Platforms

  • Design feature stores, vector stores, data ingestion frameworks, and lakehouse architectures.
  • Manage data governance, lineage, masking, and access controls around data products.

Serverless Architecture

  • Design and implement serverless solutions using AWS Lambda, API Gateway, and EventBridge.
  • Optimize serverless applications for performance, cost, and scalability.
  • Implement event-driven architectures and asynchronous processing patterns.
  • Manage serverless deployment pipelines and monitoring.

Disaster Recovery & High Availability

  • Architect and implement comprehensive disaster recovery strategies.
  • Design multi-region failover capabilities with automated recovery procedures.
  • Implement RTO/RPO requirements through backup strategies and replication.
  • Build auto-failover mechanisms using Route53 health checks and failover routing.
  • Create and maintain disaster recovery runbooks and testing procedures.
  • Ensure data durability through cross-region replication and backup strategies.

Platform Development

  • Build and maintain a self-service platform enabling rapid experimentation and testing of GenAI applications.
  • Implement Infrastructure as Code (IaC) using Terraform for consistent and repeatable deployments.
  • Create streamlined CI/CD pipelines that support local-to-dev-to-prod workflows.
  • Design systems that minimize deployment time and maximize developer productivity.
  • Establish quick feedback loops between development and deployment.

Monitoring & Operations

  • Implement comprehensive monitoring, observability, and alerting solutions.
  • Set up logging aggregation and analysis tools.
  • Ensure high availability and disaster recovery capabilities.
  • Optimize cloud costs while maintaining performance.
  • Champion DevOps best practices across the organization.
  • Automate infrastructure provisioning and application deployment.
  • Implement security best practices and compliance requirements.
  • Create documentation and runbooks for operational procedures.

Job Qualifications

Technical Skills

  • 5+ years of hands-on experience with AWS services.
  • 2+ years of hands-on experience with Databricks.
  • Expert-level knowledge of AWS core services (EC2, VPC, IAM, S3, RDS, Lambda, ECS/EKS).
  • Expert-level knowledge of Databricks capabilities.
  • Familiarity with SageMaker, Bedrock, or Anthropic/Claude API integration.
  • Strong proficiency with Terraform for infrastructure automation.
  • Demonstrated experience with containerization (Docker, Kubernetes).
  • Solid understanding of networking concepts (subnets, routing, security groups, VPN).
  • Experience with CI/CD tools (Jenkins, GitLab CI, GitHub Actions, AWS CodePipeline).
  • Proficiency in scripting languages (Python, Bash, PowerShell).

Serverless & Event-Driven Architecture

  • Extensive experience with AWS Lambda, API Gateway, ECS, Step Functions.
  • Knowledge of serverless frameworks (SAM, Serverless Framework).
  • Experience with event-driven patterns using SNS, SQS, EventBridge.
  • Understanding of serverless best practices and optimization techniques.

Disaster Recovery & Business Continuity

  • Proven experience designing and implementing DR strategies in AWS.
  • Expertise in multi-region architectures and data replication.
  • Experience with AWS backup services and cross-region failover.
  • Knowledge of RTO/RPO planning and implementation.
  • Hands-on experience with Route53 health checks and failover routing policies.
  • Primary: AWS (extensive experience required).
  • Secondary: Azure and Google Cloud Platform (working knowledge).
  • Multi-cloud architecture understanding.

Monitoring & Observability

  • Experience with monitoring tools (CloudWatch, Datadog, Prometheus, Grafana).
  • Log management systems (ELK stack, Splunk, CloudWatch Logs).
  • APM tools and distributed tracing.

Preferred Qualifications

  • AWS certifications (Solutions Architect, DevOps Engineer).
  • Databricks certifications.
  • Experience with open-source LLMs, embedding models, and RAG-based applications.
  • Experience with chaos engineering and resilience testing.
  • Knowledge of security frameworks and compliance (SOC2, HIPAA, PCI).
  • Experience implementing complex build systems for mono-repo microservices architectures.
  • Background in building developer platforms or internal tools.
  • Experience with Infrastructure as Code testing frameworks.
Seniority level
  • Seniority level
    Mid-Senior level
Employment type
  • Employment type
    Full-time
Job function
  • Job function
    Information Technology
  • Industries
    IT System Custom Software Development

Referrals increase your chances of interviewing at Power Staffing Solutions by 2x

Get notified about new Cloud Engineer jobs in Atlanta Metropolitan Area.

Software Engineer, Machine Learning (Multiple Levels) - Slack

Atlanta, GA $167,300.00-$334,600.00 1 day ago

Atlanta, GA $70,000.00-$120,000.00 1 month ago

Co-op, IT - Software Engineering (Spring, 2025)
Software Engineer Intern/Co-Op - Fall 2025

Atlanta, GA $116,000.00-$152,250.00 1 week ago

Alpharetta, GA $70,000.00-$120,000.00 12 hours ago

Associate Software Development Engineer, Crew

Atlanta, GA $120,000.00-$140,000.00 1 week ago

Alpharetta, GA $86,000.00-$125,000.00 1 month ago

Atlanta, GA $1,000.00-$2,000.00 2 months ago

Alpharetta, GA $57,200.00-$102,200.00 6 hours ago

We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Azure Data Engineer
Staff Azure Data Engineer

Synergis • Atlanta (GA)

On-site
USD 106,000 - 176,000
Medical insurance
Vision insurance
401(k)
Data Engineer
Data Engineer

Pyramid Consulting, Inc • Alpharetta (GA)

Hybrid
Health insurance
401(k) plan
Paid sick leave
Software Engineer Technical Lead
Software Engineer Technical Lead

Matlen Silver • Atlanta (GA)

On-site
USD 140,000 - 185,000
Medical insurance
Vision insurance
401(k)
Sr. Data/SW Engineer
Sr. Data/SW Engineer

Jobs via Dice • Alpharetta (GA)

Hybrid
USD 138,000 - 173,000
Senior Software Engineer
Senior Software Engineer

ATG, LLC • Atlanta (GA)

On-site
USD 130,000 - 190,000
Medical, Dental, and Vision Insurance
Health Savings Accounts/Flexible Spending Accounts
401(k) Plan with Company Match
+4
Software Engineer – Financial Applications
Software Engineer – Financial Applications

CRG • Atlanta (GA)

Remote
USD 70,000 - 120,000
Staff AI Platform Engineer — Full-Stack Cloud
Staff AI Platform Engineer — Full-Stack Cloud

Cloudera • Atlanta (GA)

On-site
USD 180,000 - 260,000
Generous PTO Policy
Flexible WFH Policy
Access to Continued Career Development
+2
Sr. Software Architect (Enterprise Systems)
Sr. Software Architect (Enterprise Systems)

SOLTECH • Duluth (GA)

Hybrid
USD 122,000 - 218,000
Medical insurance
Vision insurance
401(k)
Software Engineer - Go, Mid-Level
Software Engineer - Go, Mid-Level

Jobright.ai • Atlanta (GA)

On-site
USD 75,000 - 100,000
Medical insurance
Vision insurance
401(k)
Development Manager, Java
Development Manager, Java

Agile Resources, Inc. • Atlanta (GA)

On-site
USD 120,000 - 140,000
Unlimited PTO
Full healthcare coverage for employees + family
Disability insurance
+1