Artificial Intelligence Engineer

Sirius.

City of Melbourne

On-site

AUD 180,000 - 240,000

Full time

Just now
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Sirius. is seeking a Senior AI Engineer to design, build, and operate AI agents and tooling across the enterprise. You will ship production‑grade agents and the rollout machinery, establishing governance, versioning, and observability to ensure reliable deployments.

The role emphasizes building scalable pipelines, RAG, data platforms, and tying AI outputs to business value while maintaining security and cost discipline. A strong ML/AI foundation and hands‑on engineering are essential.

Qualifications

  • 3+ years’ experience in a fast‑paced, large‑scale commercial or enterprise environment.
  • Proven experience designing, building, and shipping LLM/agent applications to production (not just prototypes).

Responsibilities

  • Operate in an agent‑led delivery model where AI agents are trusted to do meaningful work under human supervision with emphasis on outcomes, quality controls, and repeatability.
  • Reengineer the process before you automate it, document what you build, and feed reusable patterns and guardrails back into our enterprise AI standards.
  • Help move the organisation from pilots to a governed, compounding portfolio of agents that genuinely add value—no fluff, no hype, no rework.
  • Design and ship production AI agents across taxonomy: conversational/user‑facing, knowledge & Q&A, workflow automation, decision‑support, and monitoring.
  • Reengineer and optimise the underlying business process first (mapping it in BPMN), then automate it—because automating a broken process speeds up waste.
  • Integrate tools and MCP connectors with least‑privilege, allow‑listed access, and clean separation between readers and actors.
  • Build retrieval‑augmented generation (RAG), vector, and data pipelines grounded in governed data.
  • Contribute to dashboards and insights from the governed semantic layer on Databricks with evaluation, cost monitoring, and release controls.
  • Data science and ML experience is a strong plus for forecasting, segmentation, recommendation, decision‑support, or user intelligence use cases.
  • Establish evaluation harness: quality, groundedness, safety, bias, and prompt‑injection testing with human review.
  • Versioning and CI/CD for agents, prompts, tools, MCPs, and pipelines with environments, release gates, rollback, and kill‑switches.
  • Stand up observability and AgentOps: telemetry, monitoring, cost tracking, drift detection, continuous improvement.
  • Define repeatable, supervised delivery patterns and embed them into day‑to‑day engineering and operational practice.
  • Engineer for Trust, Cost, and Scale: security, privacy, governance by design; full logging and traceability.
  • Manage performance, reliability, and cost in production; reduce recurring issues with standard patterns.
  • Document runbooks, agent cards, decision records; ensure supportability and reusability.
  • Partner with business AI champions; embed to deliver the first agent and hand it back.
  • Contribute SME insights and delivery learnings into enterprise AI standards and guardrails.

Skills

Python/TypeScript
LLM/agent apps
API versioning
Agent frameworks
MCP
RAG & embeddings
Vector databases
SQL & data pipelines
LLMOps/AgentOps
Prompt engineering
Analytical skills
Problem solving
Organisation
Adaptability

Job description

The Senior AI Engineer (Agents & Automation) designs, builds, evaluates, and operates the AI agents, tools, and pipelines that put AI to work across the business, with digital platforms as the primary focus, and the whole enterprise as the canvas. This is a hands‑on engineering role at the heart of our AI Centre of Excellence.

About the Role

You will ship production‑grade AI agents and, just as importantly, build the enterprise‑grade rollout machinery around them: a scalable, evaluated, versioned, and governed way of releasing agents, prompts, tools, MCP connectors, RAG, and data pipelines. You care as much about how reliably and repeatably we ship and operate AI as about the agents themselves.

Responsibilities
  • You will operate in an agent‑led delivery model where AI agents are trusted to do meaningful work under human supervision with the emphasis on outcomes, quality controls, and repeatability rather than approving every step.
  • You'll reengineer the process before you automate it, document what you build, and feed reusable patterns and guardrails back into our enterprise AI standards.
  • In this newly created role, you will help move the organisation from promising pilots to a governed, compounding portfolio of agents that genuinely add value—no fluff, no hype, no rework.
  • Get Excited to build agents that matter: Design and ship production AI agents across the taxonomy: conversational/user‑facing, knowledge & Q&A, workflow/process automation, decision‑support, and monitoring.
  • Reengineer and optimise the underlying business process first (mapping it in BPMN), then automate it—because automating a broken process just makes waste faster.
  • Integrate tools and MCP connectors with least‑privilege, allow‑listed access, and clean separation between agents that read and agents that act.
  • Build retrieval‑augmented generation (RAG), vector, and data pipelines grounded in governed data, so agents answer from a single source of truth.
  • Contribute to creating agents that generate dashboards, answer business questions, and release insights from enterprise data and the governed semantic layer on Databricks, with evaluation, cost monitoring, and release controls built into the delivery process.
  • Note: Data science and machine learning experience is a strong plus, particularly where it supports forecasting, segmentation, recommendation, decision‑support, or user intelligence use cases.
  • Make the Rollout Enterprise: Build the evaluation harness: quality, groundedness, safety, bias, and prompt‑injection testing, combining LLM‑as‑judge with human review, so agents earn their release.
  • Establish versioning and CI/CD for everything—agents, prompts, tools, MCPs, and pipelines with environments, release gates, rollback, and kill‑switches.
  • Stand up observability and AgentOps: telemetry, monitoring, cost tracking, drift detection, and continuous improvement loops in production.
  • Define repeatable, supervised delivery patterns and progressively embed them into day‑to‑day engineering and operational practice.
  • Engineer for Trust, Cost, and Scale: Build with security, privacy, and governance by design—human‑in‑the‑loop by default for anything touching end‑users, financials, or systems of record; full logging and traceability.
  • Manage performance, reliability, and cost in production; reduce recurring issues through standard patterns.
  • Document thoroughly—runbooks, agent cards, decision records—so what you build is supportable and reusable.
  • Raise the Whole Organisation's Game: Partner with and coach business AI champions; where a high‑value area is stuck, embed temporarily to deliver the first agent and hand it back.
  • Contribute SME insight, reusable patterns, and delivery learnings into our enterprise AI standards and guardrails through established architecture and governance forums.
Qualifications
  • 3+ years’ experience in a fast‑paced, large‑scale commercial or enterprise environment.
  • Proven experience designing, building, and shipping LLM/agent applications to production (not just prototypes).
Required Skills
  • Strong software engineering fundamentals in Python and/or TypeScript (or similar), with solid API and version‑control practice.
  • Hands‑on with agent frameworks and orchestration, tool/function calling, and the Model Context Protocol (MCP).
  • RAG, embeddings, and vector databases, plus strong SQL and data pipeline skills; comfortable working with structured and unstructured data.
  • Evaluation and LLMOps / AgentOps: building eval harnesses, versioning prompts and agents, CI/CD, observability, and guardrails for AI systems at scale.
  • Prompt and context engineering, with an eval‑driven, iterative approach.
  • Strong analytical and problem‑solving skills.
  • Results‑driven with excellent organisation and prioritisation abilities.
  • Calm, adaptable, and effective in fast‑paced environments.
  • Proactive mindset with strong initiative and accountability.
  • Commercially astute with sound business judgement.
  • Collaborative communicator, able to engage technical and non‑technical stakeholders.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Software Engineering Specialist
AI Software Engineering Specialist

Accenture • Newcastle City Council

Hybrid
AUD 170,000 - 210,000
Vendor fellowship access
Forward Deployed Engineer programme
Enterprise AI & Data Architect
Enterprise AI & Data Architect

Mane Consulting • Sydney

On-site
AUD 150,000 - 200,000
Senior Applied AI Engineer
Senior Applied AI Engineer

Rashi Joshi • Sydney

On-site
AUD 180,000 - 260,000
AI Engineer
AI Engineer

Preacta • Council of the City of Sydney

On-site
AUD 120,000 - 150,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Novus Pathway • Sydney

Hybrid
AUD 130,000 - 180,000
AI First Business Analyst Brisbane Australia
AI First Business Analyst Brisbane Australia

Phoenix DX • City of Brisbane

Hybrid
AUD 110,000 - 160,000
Enterprise Data & AI Platform Lead
Enterprise Data & AI Platform Lead

Showtime Consulting • Sydney

On-site
AUD 180,000 - 240,000
AI Engineer - Production Agents (Hybrid Sydney)
AI Engineer - Production Agents (Hybrid Sydney)

Novus Pathway • Sydney

Hybrid
AUD 130,000 - 180,000
Staff Software Engineer - AI
Staff Software Engineer - AI

Compare Club • Council of the City of Sydney

Hybrid
AUD 140,000 - 210,000
Hybrid working model
Full stack AI Engineer (Java background)
Full stack AI Engineer (Java background)

Mane Consulting • Council of the City of Sydney

On-site
AUD 250,000 - 420,000