About the Role
We are looking for a Senior Platform Infrastructure Engineer who thrives in a highly autonomous environment and enjoys taking ownership of complex technical challenges.
Our Platform Infrastructure team is responsible for the foundation that powers all production services. While the team's structure continues to evolve as we grow, its core responsibilities are well established. You'll have the opportunity to influence technical direction, improve platform reliability, and help shape future infrastructure initiatives.
This role is best suited for engineers who enjoy solving problems across multiple infrastructure domains rather than focusing on a single specialization.
What\'s You\'ll Do
- Design, build, and operate scalable infrastructure and platform services.
- Improve platform reliability through monitoring, alerting, incident management, and operational excellence.
- Develop and maintain our Kubernetes-based compute platform.
- Work with AWS production infrastructure, including compute, networking, and load balancing.
- Build and enhance observability solutions (monitoring, logging, tracing, alerting).
- Participate in security and compliance initiatives, including vulnerability remediation and infrastructure hardening.
- Support onboarding of new services and facilitate their transition between engineering teams.
- Own projects end-to-end—from identifying dependencies and coordinating with stakeholders to successful delivery.
- Drive improvements that reduce operational overhead through automation and better platform design.
- Participate in production releases and on-call rotations.
- Engineers may have opportunities to move between infrastructure teams based on business needs and personal interests, allowing exposure to different technical domains over time.
What We\'re Looking For
- Extensive experience as a Senior Platform Engineer, Infrastructure Engineer, or Site Reliability Engineer.
- Strong sense of ownership and ability to work independently with minimal supervision.
- Ability to identify problems, define priorities, and drive initiatives without waiting for detailed instructions.
- Broad infrastructure background with experience across multiple domains rather than deep specialization in only one area.
- Strong understanding of Linux internals, including: Virtual memory/Process resource management/cgroups
- Hands-on experience with: Kubernetes, Distributed systems
- AWS production infrastructure (EC2, VPC, Load Balancers, etc.)
- Networking fundamentals
- Proven ability to deliver complex technical initiatives from concept to production.
Nice to Have
Experience with one or more of the following:
- Infrastructure security and compliance
- Vulnerability management and security hardening
- Identity and access management
- Firewall configuration and privilege management
- Compliance frameworks and interpreting security requirements
- Observability platforms, monitoring, logging, and alerting
- Database infrastructure and performance troubleshooting
- CI/CD pipelines and deployment automation
- Release engineering
What Makes You Successful
- Works effectively without micromanagement.
- Takes responsibility for technical decisions and can clearly justify design choices.
- Pays close attention to detail, especially when working on security-sensitive systems.
- Thinks beyond immediate fixes and strives to reduce long-term operational burden.
- Collaborates effectively across engineering teams to unblock dependencies and deliver results.
- Is comfortable working across multiple areas of infrastructure and continuously learning new technologies.
On-Call Rotation
- The team participates in an on-call rotation to support production systems.
- Day shift rotation: approximately one week every three months.
- On-call responsibilities are closely tied to production releases, as most incidents occur during deployment activities.
- Optional night shifts are available with additional compensation.
- To maintain platform stability, releases are restricted before major holidays and during designated release freeze periods.
Why Join Us?
- Work on infrastructure that powers production at scale.
- Influence architectural decisions and long-term platform strategy.
- Solve complex engineering challenges across infrastructure, Kubernetes, reliability, security, and cloud platforms.
- Enjoy a high level of autonomy, trust, and technical ownership.
- Grow across different infrastructure domains as the platform and team continue to evolve.