Get more replies from employers
Send a job-specific resume in minutes.
V2 Solutions is seeking a Platform & Cloud Engineer to design and operate cloud-native platforms, building reusable infrastructure and owning Kubernetes clusters. The role emphasizes reliability, automation, and scalable design within a DevOps, SRE-oriented culture.
You will drive SLI/SLO definitions, participate in on-call rotations, and collaborate with software engineers to embed reliability early in the lifecycle, leveraging GitOps practices and modern observability tooling.
Role & responsibilities
Platform & Cloud Engineering Design and operate cloud-native platforms that abstract complexity and improve developer productivity
Build reusable, opinionated infrastructure using Infrastructure-as-Code
Own Kubernetes clusters, networking, service orchestration, and workload reliability
Reliability Engineering Define and drive SLIs, SLOs, and error budgets for business-critical services Participate in on-call rotations, lead incident response, and write clear, blameless RCAs Continuously reduce operational toil through automation and engineering solutions CI/CD & DevOps Build and evolve secure, automated CI/CD pipelines using GitOps principles Enable safe, frequent production deployments with strong rollback and observability Partner with application teams to embed reliability and operational excellence early in the lifecycle
Observability & Operations Implement best-in-class logging, metrics, tracing, and alerting Ensure alerts are actionable and aligned with service health not noise Build dashboards, runbooks, and self-healing mechanisms to improve MTTR.
Architecture & Collaboration Work closely with software engineers, architects, and security teams to influence system design Review infrastructure and architecture through the lens of scale, resilience, and cost efficiency Champion DevOps, SRE, and cloud-native best practices across the organization.
Preferred candidate profile
Core Engineering Strong foundations in Linux, networking (DNS, TCP/IP), and distributed systems Proficiency in Python or Go for automation and tooling Clean Git practices and strong software-engineering discipline Cloud & Containers Hands-on experience with AWS (primary); exposure to GCP or Azure is a plus Required strong experience operating Kubernetes in production environments Required experience with Helm and containerized workloads at scale. Infrastructure & Tooling Required Infrastructure-as-Code using Terraform Preferred configuration and automation using Chef / Ansible CI/CD & GitOps: ArgoCD, GitHub Actions, Jenkins, or GitLab CI. Observability & Reliability Metrics and alerting: Prometheus, Grafana, Alertmanager Tracing / APM: Datadog, New Relic, OpenTelemetry Incident management experience (PagerDuty or equivalent) Data & Messaging (Working Knowledge) Datastores: PostgreSQL, MySQL, MongoDB Streaming and search: Kafka, Elasticsearch Caching: Redis.