Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Epsilon ASI Corp. seeks a Senior Platform Engineer to design, build, and scale core cloud infrastructure and platform systems. This hands-on role intersects cloud architecture, reliability, automation, and security for a fast-growing organization.
You will implement scalable infrastructure, manage Terraform-based environments, and drive CI/CD with GitHub Actions, while embedding compliance with NIST and related standards into all workflows.
Our work starts with the real artifacts of engineering: architecture notes, pull requests, runbooks, dashboards, incidents, and the constraints teams face every day. We make that context visible, then help turn it into systems that are clearer, stronger, and easier to operate.
We work from artifacts
Architecture notes, pull requests, runbooks, dashboards, incidents, and the actual constraints your team works inside.
We make context visible
The block gives the page more texture without making it feel like a stock-photo agency site.
Embark on a rewarding journey with us. Find opportunities to grow, learn and make a lasting impact.
Join a team that embraces forward-thinking ideas, fosters innovation, and cultivates an environment where your creativity can flourish.
We think deeply about the real-world impact of every solution on teams, customers, and stakeholders.
Our work is grounded in best practices, thoughtful design, and sustainable engineering.
We're invested in outcomes that endure, not quick fixes that falter.
The interview path should feel like the work: clear communication, systems thinking, technical judgment, and the ability to collaborate without ego.
What working here should feel like
The benefits are designed around focus, trust, craft, and the reality that deep engineering work needs room.
Room for architecture, implementation, writing, review, and careful technical judgment.
A team that values context, humility, strong opinions, and better systems.
Work directly on the platform constraints that are slowing real engineering teams down.
Craft and learning
Kubernetes, cloud, modernization, AI workflows, and delivery systems in production contexts.
We care about secrets, access, change safety, auditability, and operational guardrails.
Modern tools
Use automation and AI carefully, with bounded context and human approval.
No mystery process, no puzzle interviews.
Be a part of a winning culture that fosters collaboration, creativity, and success in every career path
Full time
Hybrid
We are seeking a Senior Platform Engineer to drive the design, development, and scalability of our core cloud infrastructure and platform systems. This is a hands-on role at the intersection of cloud architecture, reliability, automation, and security. You will play a key part in building and maintaining robust, secure, and compliant environments, ensuring our systems meet the needs of a dynamic, fast-growing organization.
Architect, implement, and optimize scalable cloud infrastructure solutions.
Deploy and manage infrastructure using Terraform and infrastructure-as-code best practices.
Automate operational processes, monitoring, and deployments with modern CI/CD pipelines (e.g., GitHub Actions).
Identify, remediate, and prevent security vulnerabilities across cloud and application stacks, including containers.
Collaborate with engineering teams to integrate security, compliance, and reliability into the software development lifecycle.
Evaluate and improve cloud security controls aligned with NIST and other relevant compliance frameworks.
Monitor and troubleshoot cloud infrastructure, applications, and networking components for performance and reliability.
Contribute to technical decisions around architecture, tool selection, and infrastructure improvements.
Support continuous improvement of security posture and compliance programs.
3+ years of professional experience in platform engineering or DevOps roles.
Proven professional experience with Terraform and infrastructure-as-code.
Hands-on experience designing and managing cloud infrastructure (preferably AWS).
Strong foundation in cloud security, vulnerability remediation, and infrastructure reliability.
Experience building, maintaining, and optimizing CI/CD automation.
Solid programming or scripting ability (e.g., Go, Python, Bash).
Experience identifying and resolving container and cloud infrastructure vulnerabilities.
Familiarity with compliance standards (e.g., NIST, SOC 2, FedRAMP) and their technical requirements.
Strong analytical, troubleshooting, and problem-solving skills.
Ability to work independently and in cross-functional, remote engineering teams.
Experience managing or maintaining regulated/FedRAMP-compliant environments.
Direct experience with AWS security services and controls.
Experience with Kubernetes and container security.
Experience with vulnerability management and automated security testing.
Experience supporting government or highly regulated clients.
We're looking for candidates who can architect, troubleshoot, and operationalize core infrastructure in cloud environments -balancing reliability, compliance, and performance. You will thrive if you have experience going beyond compliance checkboxes to implement and scale robust, secure systems, and are eager to take on increasing technical and organizational responsibility as our organization grows.
This role offers considerable growth potential. As our environment and engineering teams scale, this position provides a unique path toward ongoing technical leadership and broader architecture responsibilities.
OpenTelemetry is best understood as a standard telemetry pipeline: APIs and SDKs create signals, context propagation links work across services, semantic conventions make data consistent, OTLP transports it, Collectors process it, and exporters deliver it to observability backends.
Coding-agent cost is not mainly the price of one clever prompt. It is the recurring cost of moving repository state, tool output, and loop history through paid models until useful work is accepted. Gateway observability makes that spend attributable and governable, while agent-loop discipline determines how much context gets sent.
Tool-using AI agents need more than prompt guidance. If an action can create a real side effect, enforcement should live in executable policy that can allow, deny, stop, or elevate before the tool call happens.
A practical way to distinguish DevOps, Platform Engineering, and SRE by responsibility instead of buzzword: collaboration, paved roads, and explicit reliability ownership.
Discover content that will transform your engineering organization.