Senior/Staff Infrastructure & Platform Engineer (Bay Area)

Cerebras

Santa Clara (CA)

On-site

USD 155,000 - 230,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Fortanix is seeking a Senior/Staff Infrastructure & Platform Engineer to architect, build, and evolve the infrastructure and tooling behind our products and engineering organization. This hands-on role spans cloud, on‑prem, data center, Kubernetes, networking, CI/CD and software development.

You will lead design through implementation and drive scalable, automated solutions for production environments. You’ll work across internal infra, production environments, and customer deployments,

Qualifications

  • 6+ years working in infrastructure, platform engineering, SRE, DevOps or related fields.
  • Deep Kubernetes experience including troubleshooting complex environments.
  • Experience with cloud infra (AWS/Azure/GCP) and/or on‑premises, bare‑metal or data centers.
  • Infrastructure as Code using Terraform/Ansible or similar tools.
  • Experience designing and improving CI/CD platforms and automation.
  • Strong Linux, networking, distributed systems, and infra architecture knowledge.
  • Hands‑on with Ansible, Chef, or similar automation tools.
  • Programming skills in Go, Python, Rust, or C++ for production tooling.
  • Experience with observability, monitoring, logging, and reliability practices.

Responsibilities

  • Design, build, and operate scalable, reliable infrastructure across cloud, Kubernetes, on‑premises, and hybrid environments.
  • Lead complex infra initiatives and drive architecture from problem to production.
  • Develop internal developer platforms, tools, and services to streamline development and deployment.
  • Improve CI/CD platforms and pipelines for build, test, and feedback cycles.
  • Automate provisioning, configuration, deployment, testing, monitoring, and operations workflows.
  • Enhance developer experience and remove bottlenecks across the software lifecycle.
  • Create reusable IaC, automation, and platform capabilities with scalable practices.
  • Build systems with reliability, observability, security, and disaster recovery in mind.
  • Troubleshoot complex issues, participate in incident response, and drive long‑term fixes.
  • Evaluate tech choices and set engineering standards and best practices.

Skills

6+ yrs exp
Kubernetes
AWS/Azure/GCP
Terraform/Ansible
CI/CD
Linux
Networking
Go/Python/Rust/C++
Observability

Tools

Terraform
Ansible
Chef

Job description

About Fortanix:

In today's world, where data spreads across various clouds and devices, traditional security measures aren't enough. Businesses need a dynamic approach to defend against constant cyber threats and ensure agile data security. Fortanix leads the way in data-centric cybersecurity for hybrid multicloud environments, using advanced cryptography, encryption, and confidential AI solutions.

As data breaches become more frequent and traditional defenses fall short, we focus on data exposure management to keep your information safe. Our unified data security platform addresses vulnerabilities in hybrid multicloud environments, defends against threats, and makes it easier to discover, assess, and fix data exposure risks. Whether implementing a Zero Trust model or preparing for the post-quantum computing era, we help businesses worldwide protect their most sensitive data, wherever it is.

Our commitment to solving the world’s toughest data security challenges has earned Fortanix multiple Cybersecurity Excellence and Innovation Awards, as well as recognition from industry giants such as NVIDIA, Microsoft, Intel, ServiceNow, and Snowflake.

Our team includes industry leaders and cryptography experts, creating a culture of trust, innovation and collaboration where every voice is valued. Recognized as a Great Place to Work, we're looking for passionate individuals to help us shape the future of data security and work towards a safer digital future.

About the role:

We are looking for a Senior/Staff Infrastructure & Platform Engineer to help architect, build, and evolve the infrastructure and tooling behind our products and engineering organization.

This is a highly technical, hands-on role for an engineer who can work across cloud, on-premises, data center, Kubernetes, networking, CI/CD, and software development. You will help define how our infrastructure should be designed—not simply implement predefined solutions.

You’ll work across internal infrastructure, production environments, and customer deployments, solving complex infrastructure problems and building scalable, automated solutions. The ideal candidate has a broad understanding of infrastructure architecture, deep expertise in Kubernetes and Linux, strong programming skills, and the technical leadership to take ambiguous problem statements and drive them from design through implementation.

What You’ll Do
  • Design, build, and operate scalable, reliable, and secure infrastructure across cloud, Kubernetes, on-premises, and hybrid environments.
  • Identify complex infrastructure and engineering productivity challenges and drive solutions from problem definition through architecture, implementation, and production.
  • Build and improve internal developer platforms, tools, and services that simplify development, deployment, and operational workflows.
  • Design, build, and optimize CI/CD platforms and pipelines, improving build and test performance, reliability, scalability, and developer feedback cycles.
  • Automate provisioning, configuration, deployment, testing, monitoring, and operational workflows to reduce engineering toil and improve efficiency.
  • Improve the developer experience across the software lifecycle, partnering with engineering teams to identify bottlenecks and deliver high-impact improvements.
  • Develop reusable Infrastructure as Code, automation, and platform capabilities that establish consistent and scalable engineering practices.
  • Build systems with strong reliability, observability, resilience, security, and disaster recovery capabilities.
  • Troubleshoot complex infrastructure and platform issues, participate in incident response and root-cause analysis, and drive long-term solutions to systemic problems.
  • Evaluate technologies and architectural approaches, make pragmatic technical recommendations, and establish infrastructure and engineering standards and best practices.
  • Lead technical initiatives spanning multiple engineering teams, influence architecture and engineering practices, and contribute to the long-term infrastructure and developer productivity strategy.
  • 6+ years of experience in infrastructure, platform engineering, software engineering, SRE, DevOps, or related fields.
  • Strong programming and software engineering fundamentals with experience developing production-quality tools and services.
  • Deep production Kubernetes experience, including troubleshooting complex environments.
  • Experience with cloud infrastructure such as AWS, Azure, or GCP and/or on-premises infrastructure; self-managed or bare-metal.
  • Experience with Infrastructure as Code such as Terraform, Ansible, or similar technologies.
  • Experience designing and improving CI/CD platforms and deployment automation.
  • Strong understanding of Linux, networking, distributed systems, and infrastructure architecture.
  • Strong experience with server provisioning, configuration management, patching, upgrades, and lifecycle management.
  • Hands-on experience with Ansible, Chef, or similar configuration management and automation technologies.
  • Strong programming skills in at least one language such as Go, Python, Rust, or C++.
  • Experience with observability, monitoring, logging, and reliability engineering practices.
  • Ability to lead complex technical initiatives, evaluate tradeoffs, and drive scalable, maintainable solutions through production.
  • Strong communication and collaboration skills across technical and cross-functional teams.
Nice to Have
  • Experience building internal developer platforms or developer productivity tooling.
  • Experience operating Kubernetes in on-premises, private cloud, or bare-metal environments.
  • Experience with Kubernetes operators, controllers, or Kubernetes-native tooling.
  • Experience defining engineering productivity or platform reliability metrics.
  • Kubernetes CNI/CSI and networking/storage expertise.
  • Experience with distributed technologies such as Cassandra, etcd, Ceph, Kafka, or Elasticsearch/OpenSearch.
  • Jenkins architecture or administration experience.
  • Experience with VPN, SSO, infrastructure security, or network infrastructure.
  • Hardware/server provisioning and data center operations experience.
  • Enterprise customer deployment or customer-facing infrastructure experience.
What Success Looks Like
  • You solve complex infrastructure problems across multiple layers of the technology stack.
  • You can move seamlessly between architecture, software development, automation, and hands‑on troubleshooting.
  • Improve the reliability and scalability of shared infrastructure and CI/CD systems.
  • You build reusable solutions that improve reliability, scalability, and engineering productivity.
  • Proactively identify systemic technical problems before they become significant business or operational issues.
  • You understand how Kubernetes and the infrastructure beneath it work and use that knowledge to build reliable, scalable systems.
  • We offer a collaborative work environment, amazing equity, great benefits, competitive salary, and the opportunity to redefine cloud computing.
  • Unlimited PTO (it’s between you and your work!)
  • 40 hours of Volunteer Time Off/year
  • Internet stipend
  • Friendly culture that brings the best out of everybody
  • 401k

Base salary offers for this position may vary based on factors such as location, skills, and relevant experience. Some positions may include additional compensation in the form of bonus, equity or commissions. We offer the following benefits: Medical, Dental, Vision, Life Insurance, Retirement Savings, Wellness Program, Short-and Long-Term Disability, Holidays, and more. The compensation for this role is $155,000 - $230,000 / Year.

Candidates must be legally authorized to work in the United States at the time of hire.

For this role, candidates must have a minimum of 24 months of current U.S. work authorization remaining without the need for employer sponsorship.

We are able to support H-1B transfers for candidates already in H-1B status and may consider sponsorship for candidates currently in the United States on F-1 or J-1 status. We are not initiating new visa sponsorships for candidates who would require entry into the H-1B lottery from outside the United States.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff Infrastructure & Platform Engineer (Bay Area)
Senior/Staff Infrastructure & Platform Engineer (Bay Area)

Fortanix • Santa Clara (CA), Northern (KY)

On-site
USD 155,000 - 230,000
Unlimited PTO
Volunteer Time Off (40 hours/year)
Internet stipend
+1
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)

Fortanix • Santa Clara (CA)

On-site
USD 200,000 - 290,000
Unlimited PTO
Volunteer Time Off
Internet stipend
+1
Product Support Engineer (US based)
Product Support Engineer (US based)

Fortanix, Inc. • Northern (KY)

Hybrid
USD 70,000 - 120,000
Equity compensation
Competitive salary
401k
+5
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)

Fortanix, Inc. • Santa Clara (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Unlimited PTO
Volunteer Time Off
Internet stipend
+9
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)

JobCubby • Northern (KY)

Hybrid
USD 180,000 - 230,000
Unlimited PTO
401k
Equity
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)
Staff Software Engineer - Rust and Distributed Systems (Bay Area, hybrid)

Cerebras • Santa Clara (CA)

On-site
USD 130,000 - 170,000
Unlimited PTO
401k
Internet stipend
Business Development Representative (Eastern US based)
Business Development Representative (Eastern US based)

Fortanix, Inc. • Northern (KY)

Hybrid
USD 45,000 - 90,000
Unlimited PTO
Recharge days
Volunteer Time Off
+2
Business Development Representative (Eastern US based)
Business Development Representative (Eastern US based)

Fortanix • South Carolina

On-site
USD 45,000 - 90,000
Collaborative work environment
Unlimited PTO
Quarterly recharge days
+2
Business Development Representative
Business Development Representative

Fortanix • South Carolina

On-site
USD 45,000 - 90,000
Equity
Great benefits
Unlimited PTO
+4
Staff Software Engineer - Cloud Platform (FortiSIEM)
Staff Software Engineer - Cloud Platform (FortiSIEM)

Zoomcar • Santa Clara (CA)

On-site
USD 179,000 - 219,000
Health insurance
Dental insurance
Vision insurance
+3