DevOps Engineer

Ascentt

Plano (TX)

On-site

USD 95,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ascentt is seeking a skilled cloud infrastructure manager to develop and implement cloud solutions for the automotive and manufacturing sectors. The role entails managing cloud infrastructure on AWS and Azure, developing CI/CD pipelines, and ensuring system reliability through monitoring and incident management. Ideal candidates should have strong experience in cloud platforms and DevOps methodologies.

Qualifications

  • Extensive experience with AWS and Azure cloud platforms.
  • Proficiency in developing CI/CD pipelines using GitHub Actions.
  • Strong experience with Datadog for system monitoring.

Responsibilities

  • Design and manage cloud infrastructure on AWS and Azure.
  • Develop and maintain CI/CD pipelines using GitHub Actions.
  • Implement monitoring and alerting using Datadog.

Skills

AWS
Azure
CI/CD
Datadog
Kubernetes
Terraform
Python
Bash
Security Standards

Tools

GitHub Actions
AWS CloudFormation
Datadog

Job description

Ascentt is building cutting-edge data analytics & AI/ML solutions for global automotive and manufacturing leaders. We turn enterprise data into real-time decisions using advanced machine learning and GenAI. Our team solves hard engineering problems at scale, with real-world industry impact. We’re hiring passionate builders to shape the future of industrial intelligence.

Job Description (Summary of Responsibilities):

· Cloud Infrastructure Management: Design, implement, and manage cloud-based infrastructure on AWS and Azure, ensuring optimal scalability, performance, and security.

· CI/CD Pipeline Development: Develop and maintain CI/CD pipelines using GitHub Actions for automated code deployments and testing.

· System Monitoring and Incident Management:

· Implement and configure Datadog for comprehensive system monitoring.

· Develop and maintain Datadog dashboards to visualize system performance and metrics.

· Set up proactive alerts in Datadog to detect and respond to incidents swiftly, ensuring high system reliability and uptime.

· Conduct root cause analysis of incidents and implement corrective actions using Datadog insights.

· Collaboration with AI Teams: Work closely with AI teams to support the operational aspects of LLMs, including deployment strategies and performance tuning.

· Infrastructure as Code (IaC): Implement IaC using tools like Terraform or AWS CloudFormation to automate infrastructure provisioning and management.

· Container Orchestration: Manage container orchestration systems such as Kubernetes or AWS ECS.

· Operational Support for LLMs: Provide operational support for LLMs, focusing on performance optimization and reliability.

· Scripting and Automation: Utilize scripting languages such as Python and Bash for automation and task management.

· Security and Compliance: Ensure compliance with security standards and best practices, implementing robust security measures.

· Documentation: Document system configurations, procedures, and best practices for internal and external stakeholders.

· DevOps Collaboration: Work with development teams to optimize deployment workflows, introduce best practices for DevOps, and improve overall efficiency.

· Technology and Industry Awareness: Stay up-to-date with emerging technologies and industry trends to suggest improvements and upgrades.

Qualifications and Skills Required:

· Extensive experience with AWS and Azure cloud platforms.

· Proficiency in developing CI/CD pipelines using GitHub Actions.

· Strong experience with Datadog for system monitoring, including implementation, configuration, and maintenance.

· Demonstrated ability to create and maintain Datadog dashboards for performance visualization.

· Proven expertise in setting up alerts and conducting incident response with Datadog.

· Hands-on experience with container orchestration systems such as Kubernetes or AWS ECS.

· Proficiency in Infrastructure as Code (IaC) tools like Terraform or AWS CloudFormation.

· Familiarity with operational aspects of Large Language Models (LLMs) is highly desirable.

· Strong scripting skills in Python, Bash, or similar languages.

· In-depth knowledge of security standards and best practices.

· Proven ability to work collaboratively with development and AI teams.

· Commitment to staying current with industry trends and emerging technologies

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Platform Engineer - 90/HR- REMOTE
Data Platform Engineer - 90/HR- REMOTE

ContractStaffingRecruiters.com • Branford (CT)

Remote
USD 130,000 - 160,000
Senior Cloud Engineer
Senior Cloud Engineer

Ascentt • Plano (TX)

On-site
USD 120,000 - 150,000
Sr. Principal Data Scientist / Machine Learning Engineer
Sr. Principal Data Scientist / Machine Learning Engineer

Ascentt • Plano (TX)

On-site
USD 130,000 - 180,000
Senior Data Engineer / AI ML Engineer with Python, AI/ML & LLMs
Senior Data Engineer / AI ML Engineer with Python, AI/ML & LLMs

Ampcus Inc • Reston (VA)

On-site
USD 100,000 - 130,000
Technical Lead
Technical Lead

Luxoft • Irvine (CA)

On-site
USD 180,000 - 240,000
AI & Data Architect
AI & Data Architect

Insaito Software • United States

Remote
USD 150,000 - 230,000
Senior AI/ML Engineer
Senior AI/ML Engineer

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Software Engineer II – Enterprise AI Products
Software Engineer II – Enterprise AI Products

Jobtailor • Connecticut

On-site
USD 140,000 - 180,000
AWS Infrastructure Engineer - Senior
AWS Infrastructure Engineer - Senior

Compunnel, Inc. • Columbus (OH)

On-site
USD 100,000 - 130,000
Principal AI Engineer
Principal AI Engineer

Stellantis Financial Services • Auburn Hills (MI)

On-site
USD 180,000 - 240,000