Описание
CloudLinux builds Linux infrastructure and security products. Its Automation & Management Services cell works across teams and services on cloud-cost data, infrastructure inventory, network policies, and capacity workflows.
Задачи
- Turn incomplete briefs into agreed outcomes and workable plans
- track cross-team dependencies, explain trade-offs, raise blockers early, communicate estimate changes, and confirm completion with users
- Build maintainable software and cloud/infrastructure API integrations, handling credentials, pagination, retries, partial failures, and repeated execution to keep unattended workflows safe and reliable
- Transform billing, usage, and inventory data into repeatable outputs
- reconcile totals and periods, apply agreed allocation rules, and surface missing or uncertain mappings
- Extend inventory, network-policy, and capacity automation using existing systems
- retain necessary human decisions and verify operational results
- Deliver changes through version-controlled code and CI
- understand live state, test recovery, release safely, and diagnose failures across code, Linux, APIs, data, and networking
- Own monitoring, actionable alerts, documentation, and maintenance
- leave systems another engineer can operate and use coding agents while remaining responsible for design, code, tests, and outcomes
- The work includes new development, improvements, repairs and maintenance.
- There is no traditional on-call rotation
- you respond to alerts from your own systems during working hours.
Требования
- Have typically 5+ years of senior-level experience in software, infrastructure, platform, or automation engineering; be able to explain design decisions, implementation, and operational results
- Have strong production Python, software design, testing, debugging, and API integration skills; be able to understand unfamiliar code and improve existing systems
- Have hands‑on AWS experience and solid Linux and networking fundamentals, including diagnosis across host, service, network, and cloud boundaries
- Have practical SQL and data‑validation skills, including joins, units, periods, ownership mappings, and reconciliation; distinguish evidence from assumptions
- Have experience delivering infrastructure safely through IaC and CI, including state, drift, idempotency, dependencies, and rollback, using tools such as Terraform/OpenTofu and Ansible
- Demonstrate advanced coding‑agent workflows for real engineering tasks: plan and delegate substantial work, provide context, run agents without continuous supervision within defined permissions and stop conditions, and independently explain, debug, test, and verify their output
- Deliver autonomously and collaborate clearly; investigate problems, obtain decisions, challenge unsafe or unnecessarily complex approaches, and follow through
- Have English B2 or higher for technical discussions, written decisions, and internal customer communication
- Будет плюсом: cloud billing or FinOps tooling (e.g. AWS CUR/FOCUS, Athena), inventory or event‑driven integrations (e.g. NetBox or similar), network policies, flow monitoring, identity and access management, capacity tooling, additional cloud platforms, OpenNebula, Kubernetes, Go, self‑service systems, or experience with Python, Ansible, Terraform/OpenTofu, GitLab CI/CD, Grafana, and related observability tools; ability to quickly learn unfamiliar technologies, products, and internal business rules
Условия
- Flexible working hours
- Paid 24 days of vacation per year
- 10 days of national holidays
- unlimited sick leaves
- Compensation for private medical insurance
- Co‑working and gym/sports reimbursement
- Budget for education
- Opportunity to receive a reward for the most innovative idea that the company can patent
- A focus on professional development
- Interesting and challenging projects
- Fully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide