Build, automate and scale critical digital services.
We are seeking a skilled Site Reliability Engineer to join a high-performing technology team responsible for the reliability, security and continuous improvement of critical cloud platforms. This role combines engineering, automation and operational excellence, with a strong focus on modern AWS environments, CI/CD pipelines and platform reliability.
As an SRE, you will play a key role in designing, implementing and operating application delivery pipelines while helping to simplify, secure and modernise cloud-based services. You'll work across AWS infrastructure, Kubernetes platforms and automation tooling to drive efficiency, scalability and resilience.
What You'll Do
- Design, build and support CI/CD pipelines and deployment processes.
- Implement automation that improves operational efficiency, reliability and security.
- Support and optimise AWS-hosted platforms, websites, backend APIs and cloud services.
- Monitor system health, investigate incidents and perform root cause analysis.
- Drive improvements in platform performance, availability and scalability.
- Collaborate with developers, security specialists and stakeholders to deliver robust solutions.
- Contribute to infrastructure-as-code and cloud engineering initiatives.
- Document processes, procedures and operational standards.
You'll work with a modern cloud-native stack including:
- AWS (EKS, ECS, CloudFront, WAF, ACM, ALB, RDS, CloudWatch, S3, EC2)
- GitHub Actions
- AWS CodeBuild & CodePipeline
- Terraform
- Kubernetes
- CDK & TypeScript
- Bash scripting
- Git/GitHub
- CloudWatch and emerging observability platforms including Grafana and Prometheus.
Essential Skills & Experience
- Experience developing automation through scripting and programming.
- Strong AWS platform knowledge and cloud operations experience.
- Proven experience implementing and supporting CI/CD pipelines.
- Experience with infrastructure as code, particularly Terraform.
- Knowledge of Git and modern source control practices.
- Experience with monitoring, alerting and operational support.
- Strong troubleshooting and root cause analysis capabilities.
- Understanding of security across code, infrastructure and delivery processes.
- Excellent communication and teamwork skills.
- A commitment to continuous learning and professional development.
- Python, Go, PHP or TypeScript development experience.
- Kubernetes and container platform experience.
- Jira/Atlassian tools.
- MongoDB.
- Automated testing tools such as Playwright.
- Experience with Bitrise or similar DevOps tooling.
If you are interested and possess the right experience, please apply now via the link to be considered.
Contact: Laura GILLES – (08) 9423 1416 – (Job reference: 271548 )
Peoplebank and Leaders IT are committed to creating a diverse and inclusive workplace where everyone belongs. We welcome applications from people of all backgrounds, identities, and experiences. If you need adjustments to the recruitment process due to your circumstances, please let us know—we’re here to support you.

