Platform Engineer

Tru Staffing Inc

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

15 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tru Staffing Inc. seeks a Platform Engineer to advance our AI-enabled engineering stack from a London-based hub, with a predominantly remote setup.

You will design and operate pipelines, IaC, and cloud services across AWS/Azure and Kubernetes, aligning with GitOps and security best practices. You will implement observability, self‑service tooling, and automation while collaborating with developers to ship reliable software and AI workloads at scale.

Qualifications

  • Proficiency writing production‑quality code in Go and Python (plus PowerShell/Bash), using Git, PRs, code review, and automated testing.
  • Working proficiency building pipelines-as-code (GitHub Actions, Azure DevOps, GitLab CI, Jenkins).
  • Experience with infrastructure as code (Terraform, Ansible, Helm) and GitOps concepts.
  • Hands‑on experience with at least one major cloud platform (AWS or Azure) and Kubernetes/Docker.
  • Familiarity with observability tooling (Grafana, Datadog, Splunk, ELK, OpenTelemetry) and basic SRE practices.
  • Exposure to test automation, policy-as-code, and platform security practices.
  • Familiarity with ITIL best practices (incident, change, and problem management) preferred.
  • Experience with Lean or Agile methodologies preferred.
  • Relevant certifications (e.g., AWS/Azure Associate, Certified Kubernetes Administrator) preferred.

Responsibilities

  • Build and maintain self-service tooling, CLIs, libraries, and templates, as version‑controlled, tested code, that make it easy for teams to build, test, and ship software.
  • Contribute features and fixes to the internal developer platform and portal through pull requests and code review.
  • Provide day‑to‑day support to developers using platform services, triaging and resolving requests and issues.
  • Support governed self‑service access to AI platform capabilities, including model access, retrieval tooling, and evaluation workflows, so AI capabilities can move from prototype to production using standard platform patterns.
  • Implement and maintain pipelines-as-code (e.g., GitHub Actions/Azure DevOps YAML) to established patterns, keeping builds, tests, and deployments reliable and fast.
  • Execute and support release activities, following defined GitOps and change‑management processes.
  • Write and maintain automated tests and quality gates within pipelines.
  • Implement evaluation‑based quality gates for AI systems, including evals‑as‑code, regression suites against golden datasets, and human‑review thresholds where required.
  • Author and maintain modular, tested infrastructure-as-code (e.g., Terraform, Helm) to provision and configure cloud and on‑prem resources.
  • Deploy and operate workloads across cloud platforms (AWS/Azure) and Kubernetes using GitOps (Argo CD/Flux).
  • Follow tagging, cost, and configuration standards when provisioning resources.
  • Help deploy and operate AI workload infrastructure, including model gateways, retrieval services, orchestration components, and supporting cloud or Kubernetes resources.
  • Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry).
  • Participate in the on‑call rotation, responding to incidents and helping restore service.
  • Contribute to blameless post‑incident reviews and implement follow‑up actions in code.
  • Instrument AI services for operational visibility, including tracing across prompts, tools, and agent steps, and monitoring latency, cost, token usage, and quality regressions.
  • Apply platform security controls and remediate vulnerabilities to defined standards, including policy‑as‑code checks (e.g., OPA/Conftest).
  • Manage secrets, access, and configuration securely using approved tooling (e.g., Vault).
  • Apply security and confidentiality controls for AI workloads, including prompt‑injection and data‑exfiltration defenses, output filtering, and access controls across prompts, retrieval, and agent tool use.
  • Automate repetitive operational tasks by writing scripts and small services.
  • Follow engineering standards, version‑control workflows, code‑review, and testing practices.
  • Collaborate with platform and development teams and elevate complex issues appropriately.
  • Create and maintain runbooks, SOPs, and documentation as code alongside the tooling it describes.

Skills

Go
Python
PowerShell
Bash
Pipelines as code
Terraform
Ansible
Helm
Kubernetes
GitHub Actions
Azure DevOps
GitLab CI
Jenkins
Git
AWS
Azure
Grafana
Prometheus
Datadog
Splunk
OpenTelemetry
Policy-as-code
ITIL
Agile
Kubernetes Admin

Education

Bachelor’s degree in Computer Science, Information Technology, Engineering, or related field

Tools

Terraform
Ansible
Helm
Kubernetes
AWS
Azure
Git
GitHub Actions
Azure DevOps
GitLab CI
Jenkins
Terraform
PowerShell
Bash
Prometheus
Grafana
Datadog
OpenTelemetry

Job description

Skip to content Back Platform Engineer London| United Kingdom Direct Hire

Our client, a leading global law firm, is seeking a Platform Engineer to join its technology team in a primarily remote role, (however candidates must be within commutable distance of the London office for occasional onsite visits). This is a highly hands‑on, AI/ML-enabled engineering position focused on building the platforms and automation that power modern software and AI workloads across the firm. The successful candidate must be able to code in both Go and Python and will work extensively across pipelines-as-code, infrastructure-as-code, policy-as-code, test automation, and the operationalization of data and AI workflows. It’s an excellent opportunity for an experienced Platform/DevOps Engineer to work with cloud, Kubernetes, CI/CD, observability, and emerging AI infrastructure while helping establish scalable, secure, and reliable engineering practices. This is an opportunity to join an innovative, progressive, and collaborative team.

Primary applications and platforms include:
  • CI/CD, Source Control & Test Automation: GitHub Actions, Azure DevOps, GitLab CI, Jenkins; Git, JFrog Artifactory; Playwright, pytest/JUnit
  • Infrastructure & Config as Code: Terraform, Ansible, Bicep/ARM, Helm, Kustomize; GitOps via Argo CD and Flux
  • Cloud & Orchestration: AWS, Azure, Docker, Kubernetes
  • AI Inference & Application Infrastructure: Frontier and open-weight models via Anthropic, Azure OpenAI, and Amazon Bedrock; model gateways and routing, retrieval and hybrid search, document ingestion, tool/function calling, Model Context Protocol (MCP), and agent orchestration
  • AI Evaluation & Quality: Eval harnesses and golden datasets, LLM-as-judge and human-in-the-loop review, regression suites, and red-teaming
  • Observability & Monitoring: Prometheus, Grafana, Datadog, Splunk, Elastic/ELK, OpenTelemetry, including GenAI tracing and token, latency, and cost telemetry
  • Platform Security & Policy-as-Code: HashiCorp Vault, OPA/Conftest, SAST/DAST
  • Developer Portal & Self-Service: Internal developer portal, CLIs/SDKs, and APIs
Responsibilities include:
Developer Experience & Self-Service Enablement
  • Build and maintain self-service tooling, CLIs, libraries, and templates, as version‑controlled, tested code, that make it easy for teams to build, test, and ship software.
  • Contribute features and fixes to the internal developer platform and portal through pull requests and code review.
  • Provide day‑to‑day support to developers using platform services, triaging and resolving requests and issues.
  • Support governed self‑service access to AI platform capabilities, including model access, retrieval tooling, and evaluation workflows, so AI capabilities can move from prototype to production using standard platform patterns.
CI/CD, Release Management & Test Automation
  • Implement and maintain pipelines-as-code (e.g., GitHub Actions/Azure DevOps YAML) to established patterns, keeping builds, tests, and deployments reliable and fast.
  • Execute and support release activities, following defined GitOps and change‑management processes.
  • Write and maintain automated tests and quality gates within pipelines.
  • Implement evaluation‑based quality gates for AI systems, including evals‑as‑code, regression suites against golden datasets, and human‑review thresholds where required.
Infrastructure as Code & Cloud Platforms
  • Author and maintain modular, tested infrastructure-as-code (e.g., Terraform, Helm) to provision and configure cloud and on‑prem resources.
  • Deploy and operate workloads across cloud platforms (AWS/Azure) and Kubernetes using GitOps (Argo CD/Flux).
  • Follow tagging, cost, and configuration standards when provisioning resources.
  • Help deploy and operate AI workload infrastructure, including model gateways, retrieval services, orchestration components, and supporting cloud or Kubernetes resources.
Observability, Monitoring & Site Reliability (SRE)
  • Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry).
  • Participate in the on‑call rotation, responding to incidents and helping restore service.
  • Contribute to blameless post‑incident reviews and implement follow‑up actions in code.
  • Instrument AI services for operational visibility, including tracing across prompts, tools, and agent steps, and monitoring latency, cost, token usage, and quality regressions.
Platform Security & Automation
  • Apply platform security controls and remediate vulnerabilities to defined standards, including policy‑as‑code checks (e.g., OPA/Conftest).
  • Manage secrets, access, and configuration securely using approved tooling (e.g., Vault).
  • Apply security and confidentiality controls for AI workloads, including prompt‑injection and data‑exfiltration defenses, output filtering, and access controls across prompts, retrieval, and agent tool use.
  • Automate repetitive operational tasks by writing scripts and small services.
Software Engineering Standards & Collaboration
  • Follow engineering standards, version‑control workflows, code‑review, and testing practices.
  • Collaborate with platform and development teams and elevate complex issues appropriately.
  • Create and maintain runbooks, SOPs, and documentation as code alongside the tooling it describes.
Qualifications:
  • Proficiency writing production‑quality code in Go and Python (plus PowerShell/Bash), using Git, pull requests, code review, and automated testing.
  • Working proficiency building pipelines-as-code (GitHub Actions, Azure DevOps, GitLab CI, Jenkins).
  • Experience with infrastructure as code (Terraform, Ansible, Helm) and GitOps concepts.
  • Hands‑on experience with at least one major cloud platform (AWS or Azure) and Kubernetes/Docker.
  • Familiarity with observability tooling (Grafana, Datadog, Splunk, ELK, OpenTelemetry) and basic SRE practices.
  • Exposure to test automation, policy-as-code, and platform security practices.
  • Familiarity with ITIL best practices (incident, change, and problem management) preferred.
  • Experience with Lean or Agile methodologies preferred.
  • Relevant certifications (e.g., AWS/Azure Associate, Certified Kubernetes Administrator) preferred.
Experience:
  • Significant experience in platform, DevOps, infrastructure, or software engineering roles.
  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field (or equivalent experience).
  • Experience supporting business‑critical systems, ideally within professional services, legal, or regulated environments.

Job ID: 7604

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer
Platform Engineer

Gibson Dunn • Greater London

On-site
GBP 90,000 - 125,000
AI Platform Engineer
AI Platform Engineer

Tempest Vane Partners • Greater London

Hybrid
GBP 120,000 - 180,000
Competitive compensation
Discretionary bonus
Benefits package
Software Engineer, General
Software Engineer, General

United States Digital Space LLC • Greater London

On-site
GBP 70,000 - 110,000
Wellness benefits
Flexible work options
Competitive compensation
+1
Software Engineer (Platform)
Software Engineer (Platform)

LinuxRecruit • City Of London

Hybrid
GBP 50,000 - 75,000
High autonomy
Flexible hybrid setup
Equity options
Senior Platform Engineer
Senior Platform Engineer

Atarus • Greater London

On-site
GBP 90,000 - 120,000
Principal Platform Engineer
Principal Platform Engineer

Trayport • Greater London

On-site
GBP 120,000 - 170,000
Diversity & inclusion
Accommodations for applicants
Lead Platform Engineer (AI Early Stage Startup)
Lead Platform Engineer (AI Early Stage Startup)

Foundation Partners • Harrow

Hybrid
GBP 90,000 - 150,000
Senior ML Platform Engineer - Artificial Intelligence
Senior ML Platform Engineer - Artificial Intelligence

Bloomberg LP • Greater London

On-site
GBP 80,000 - 100,000
Platform Engineer
Platform Engineer

LinuxRecruit • Greater London

Hybrid
GBP 70,000 - 90,000
AI Platform Engineer
AI Platform Engineer

Jobtailor • Glasgow

On-site
GBP 95,000 - 120,000