A leading AI-native platform company is seeking a Founding Platform & Reliability Engineer to oversee the design and reliability of their entire infrastructure. This role demands a mix of hands-on implementation and strategic architecture decisions, ensuring system reliability and cost efficiency in a rapidly evolving environment. Key qualifications include over 5 years of experience in production systems and strong engineering skills in cloud-native environments. Enjoy competitive compensation and a hybrid work setup in the Bay Area.
Qualifications
5+ years building and operating production systems where reliability and scaling are core.
Comfortable working across cloud infrastructure and distributed systems.
Ability to operate with ambiguity and define problems before solving them.
Responsibilities
Define and operationalize SLOs/SLIs across critical user journeys.
Implement reliability patterns at external boundaries.
Act as a senior technical voice influencing architecture and best practices.
Skills
Cloud-native experience (AWS or GCP)
Strong software engineering skills
Deep knowledge of observability practices
Ability to design resilient interactions
Can communicate tradeoffs to peers
Tools
GCP
Node.js
TypeScript
Python
React / Next.js
Job description
A leading AI-native platform company is seeking a Founding Platform & Reliability Engineer to oversee the design and reliability of their entire infrastructure. This role demands a mix of hands-on implementation and strategic architecture decisions, ensuring system reliability and cost efficiency in a rapidly evolving environment. Key qualifications include over 5 years of experience in production systems and strong engineering skills in cloud-native environments. Enjoy competitive compensation and a hybrid work setup in the Bay Area.