Overview
At PwC, our people in infrastructure focus on designing and implementing robust, secure IT systems that support business operations. They enable the smooth functioning of networks, servers, and data centres to optimise performance and minimise downtime. Those in cloud operations will focus on managing and optimising cloud infrastructure and services to enable seamless operations and high availability for clients. You will be responsible for monitoring, troubleshooting, and implementing industry leading practices for cloud-based systems.
Key Skills
- Apply a learning mindset and take ownership for your own development.
- Appreciate diverse perspectives, needs, and feelings of others.
- Adopt habits to sustain high performance and develop your potential.
- Actively listen, ask questions to check understanding, and clearly express ideas.
- Seek, reflect, act on, and give feedback.
- Gather information from a range of sources to analyse facts and discern patterns.
- Commit to understanding how the business works and building commercial awareness.
- Learn and apply professional and technical standards, uphold the Firm's code of conduct and independence requirements.
Job Summary
We are seeking an Oracle database administrator to design, deploy, configure, secure, monitor, and maintain enterprise databases (SQL Server, Oracle, MySQL, PostgreSQL, etc.) across development, test, and production environments. The ideal candidate will build and manage automated provisioning and change pipelines, implement HA/DR and backup strategies, enforce security and compliance standards, and continuously monitor performance and capacity to ensure high availability and optimal operation.
Minimum Qualifications
- Bachelor’s Degree in IT, Computer Science, or a related technical field.
- 2-4 years of experience in database administration.
- Experience with Oracle, MySQL, PostgreSQL, or similar.
- Basic knowledge of Terraform, Ansible, and CI/CD tools.
- Strong communication, documentation, and collaboration skills.
Key Responsibilities
- Database Provisioning & Infrastructure-as-Code
- Deploy and configure new database instances (dev, test, prod) in accordance with system and storage requirements.
- Build and maintain automated provisioning pipelines using Terraform, Ansible, and CI/CD tools.
- Enforce infrastructure-as-code standards, audit logging, and reporting.
- Validate automation through test runs and integrate database changes into CI/CD workflows.
- Backup, Recovery & Disaster Recovery – design and schedule daily, weekly, and monthly backups; configure and validate backup jobs in native DBMS tools.
- Test recovery scenarios quarterly and maintain up-to-date DR runbooks.
- Monitor backup job success/failure; escalat[e]e and remediate missed or failed backups.
- Implement and validate log shipping, Always On Availability Groups, Oracle RAC or clustering for DR.
- Patch Management & Upgrades – plan and coordinate quarterly patch reviews and database version upgrades.
- Apply vendor-released patches, perform post-upgrade validation, and resolve any functional regressions.
- Performance Monitoring, Tuning & Capacity Planning – monitor CPU, memory, storage usage, query performance, waits and latches using native and third‑party tools.
- Run baseline health checks post‑maintenance, resolve anomalies, and document changes.
- Forecast resource growth (12–24 months), identify workload spikes, and plan infrastructure scaling.
- Recommend and implement index and query optimizations in collaboration with application owners.
- High Availability & Replication – design, configure, and validate HA solutions including clustering, Always On, log shipping, and replication.
- Monitor replication jobs, troubleshoot latency/errors, and conduct periodic failover drills.
- Security, Compliance & Audit – implement database security standards: encryption, data masking, role‑based access control, schema, and user setup.
- Approve or deny access requests, remove orphaned accounts, and audit user activity.
- Generate compliance reports, track policy adherence, and escalat[e]e any violations.
- Incident Management & Root Cause Analysis – monitor for deadlocks, transaction failures, data corruption, and alert conditions.
- Troubleshoot and resolve incidents in partnership with application teams or escalat[e]e to vendors.
- Document root‑cause analyses, corrective actions, and update runbooks.
- Reporting & Dashboarding – produce weekly performance and availability dashboards; provide insights and recommendations for tuning, optimisation, and capacity upgrades.
- Scripting & Task Automation – develop and maintain scripts in Shell, PowerShell, PL/SQL, T‑SQL, etc., and store them in version control.
- Automate routine tasks such as patching, backups, user provisioning, and environment deployments.
- Test scripts in lower environments and deploy to production via automated pipelines.
Preferred Knowledge & Skills
- Strong expertise with Terraform, Ansible, Git, Jenkins/GitLab CI or equivalent CI/CD tools.
- Deep understanding of HA/DR architectures, backup/restore processes, and disaster recovery testing.
- Proven ability to monitor, troubleshoot, and tune database performance at scale.
- Familiarity with security best practices, compliance frameworks, and audit reporting.
- Proficient scripting skills (Shell, PowerShell, PL/SQL, T‑SQL) and infrastructure‑as‑code principles.
- Excellent communication, documentation, and collaboration skills.
- Hands‑on knowledge of Oracle (RAC, Data Guard, Log Shipping), MySQL, MariaDB, MongoDB, PostgreSQL, PowerShell, AWS RDS, Aurora, NoSQL.
- Cloud infrastructure knowledge – Azure, AWS, GCP – and related tools such as Terraform, Solarwinds SQL Diagnostics, DBArtisan.
Certifications
Preferred certifications: Oracle OCA, ITIL Foundation, Oracle OCP, AWS/GCP database certifications.