Software Engineer – Platform Security

FriendliAI

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation and benefits package
Daily lunch and dinner provided
Unlimited snacks and beverages
Health check-up and top-tier hardware support
Flexible working hours

Job summary

FriendliAI is looking for a Forward Deployed Engineer (FDE) to assist enterprises in deploying and operating generative AI workloads. The role involves collaborating with customers to address AI inference challenges, manage Kubernetes workloads, and contribute to deployment strategies. Ideal candidates should have over 3 years of cloud infrastructure or DevOps experience, proficiency with crucial tools like Kubernetes and Docker, and a relevant degree. This position offers a competitive salary, flexible hours, and unique benefits including meals and health support.

Qualifications

  • 3+ years of experience in cloud infrastructure, DevOps, or reliability engineering.
  • Proficiency with Kubernetes, Docker, Terraform, and Helm.
  • Strong problem-solving and debugging skills in real-world environments.

Responsibilities

  • Design and implement large-scale deployment architectures for LLM and multimodal inference.
  • Deploy and manage containerized workloads across Kubernetes clusters.
  • Collaborate with customers’ DevOps teams to integrate FriendliAI’s infrastructure into their CI/CD workflows.

Skills

Cloud infrastructure
DevOps
Reliability engineering
Kubernetes
Docker
Terraform
Helm
Distributed systems
Networking
Performance tuning

Education

Bachelor's or Master's degree in a relevant field

Tools

Kubernetes
Docker
Terraform
Helm
AWS
GCP
OCI

Job description

About the job

FriendliAI is seeking a Forward Deployed Engineer (FDE) to assist enterprises in deploying, scaling, and operating generative and agentic AI workloads on FriendliAI infrastructure. You will work directly with customers to solve and implement production-grade applications using our products, such as Serverless Endpoints, Dedicated Endpoints, or Container.

Friendli Container is our service that allows customers to download our inference engine as Docker images and deploy it in their chosen environment, such as private clouds or on-premises. Our Friendli Container can be adopted directly to AWS EKS clusters using our EKS add-on product.

You will work directly on our customers’ projects, collaborating with their engineering teams to solve AI inference challenges like scaling, orchestration, and monitoring. This is a hands-on, customer-embedded role. If you have worked in DevOps, platform engineering, or SRE for AI applications, this is your ideal position.

Key Responsibilities
  • Design and implement large-scale deployment architectures for LLM and multimodal inference
  • Deploy and manage containerized workloads across Kubernetes clusters
  • Diagnose production issues, such as performance bottlenecks, and implement temporary fixes as needed
  • Collaborate with customers’ DevOps teams to integrate FriendliAI’s infrastructure into their CI/CD workflows
  • Develop scripts, Helm charts, and Terraform modules that simplify repeated deployments
  • Contribute field insights to shape our platform reliability, observability, and scaling strategies
  • Lead workshops, technical sessions, or webinars to help customers master infrastructure best practices
Qualifications
  • 3+ years of experience in cloud infrastructure, DevOps, or reliability engineering
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent
  • Proficiency with Kubernetes, Docker, Terraform, and Helm
  • Strong foundation in distributed systems, networking, and performance tuning
  • Experience with GPU-based computing and generative AI model serving workloads
  • Strong technical background in backend systems or AI tooling
  • Experience operating workloads on AWS, GCP, or OCI
  • Excellent problem-solving and debugging skills in real-world environments
Preferred Experience
  • Experience deploying large models (LLMs, diffusion models) on GPUs or clusters
  • Familiarity with inference frameworks (Triton, vLLM, TensorRT, DeepSpeed-Inference)
  • Familiarity with observability stacks (Prometheus, Grafana, Loki, ELK, OTEL)
  • Understanding of networking security and compliance frameworks (e.g., SOC 2)
  • Experience supporting on-prem or hybrid-cloud deployments
Benefits
  • A front-row seat to the generative AI infrastructure revolution
  • Competitive compensation and benefits package
  • Daily lunch and dinner provided; unlimited snacks and beverages
  • Health check-up and top-tier hardware support
  • Flexible working hours and a highly collaborative environment
About us

FriendliAI is building the next-generation AI inference platform that accelerates the deployment of large language and multimodal models with unmatched performance and efficiency. Our infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 510,000 open-source models. We are on a mission to deliver the world’s best platform for AI inference.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Solutions Architect - AI Inference Specialist
Solutions Architect - AI Inference Specialist

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive compensation and benefits package
Daily lunch and dinner
Unlimited snacks and beverages
+2
Software Engineer - Full Stack
Software Engineer - Full Stack

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided
Health check-up support
+2
Solutions Architect - AI Model Specialist
Solutions Architect - AI Model Specialist

FriendliAI • San Francisco (CA)

On-site
USD 110,000 - 150,000
Daily lunch and dinner provided
Unlimited snacks and beverages
Health check-up
+2
Software Engineer – Cloud Infrastructure
Software Engineer – Cloud Infrastructure

FriendliAI • San Francisco (CA)

On-site
USD 150,000 - 190,000
Flexible working hours
Lunch and dinner provided
Health check-up support with top-tier硬
+1
Software Engineer – AI Agents
Software Engineer – AI Agents

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Flexible working hours
Daily lunch and dinner
Health check-up support
+1
Account Executive
Account Executive

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 150,000
Highly competitive base salary plus uncapped commission
Equity in a growing AI company
Comprehensive benefits package
+1
Forward Deployed Engineer
Forward Deployed Engineer

Eliza • United States

On-site
USD 90,000 - 120,000
Competitive compensation
Equity options
Flexible work across industries
Senior Software Engineer - Infrastructure / Platform (AI Startup)
Senior Software Engineer - Infrastructure / Platform (AI Startup)

Mintstage • California (MO)

On-site
USD 140,000 - 200,000
Software Engineer – AI Inference Engine
Software Engineer – AI Inference Engine

FriendliAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Flexible working hours
Daily lunch and dinner provided; unlimited snacks and beverages
Health check-up support and top-tier equipment/hardware support
+2
Senior AI Engineer - Forward Deployed (FDE)
Senior AI Engineer - Forward Deployed (FDE)

Moring AI • Atlanta (GA)

On-site
USD 150,000 - 210,000
Travel up to 30%