- 5+ years in infrastructure, platform engineering, SRE or related operational role.
- Hands-on experience running and troubleshooting Kubernetes in production, ideally AWS EKS.
- Practical AWS (VPCs, IAM, EC2, DNS/networking) and experience managing infrastructure via Terraform.
Role Type
Pay Rate
Description
Xestro is growing rapidly, and we’re looking for an experienced Platform Engineer to help operate and improve our production infrastructure based in Maroochydore.
About Xestro
Xestro is an Australian healthcare software company that helps specialist doctors manage their patients and run their practices. Our cloud-based platform is used daily in hundreds of busy clinics across Australia by thousands of healthcare providers.
We’re based in Maroochydore on the Sunshine Coast, close to the Sunshine Plaza, and have recently moved into a multi-storey office to support our next phase of growth.
Our platform runs on AWS, with an established Linux, Apache, PHP and MySQL environment alongside a growing Kubernetes-based platform. We’re focused on improving reliability, security and automation while modernising and scaling our infrastructure.
The Role
You’ll report to the Head of Engineering and take day-to-day technical direction from our Lead SRE. You’ll join our Platform team, working hands-on with Kubernetes, AWS EKS, Terraform and deployment pipelines to support a production system used every day in real healthcare settings.
The work includes operating production clusters, troubleshooting infrastructure and deployment issues, managing upgrades, monitoring performance and capacity, and improving automation. You’ll work with GitLab, ArgoCD, Temporal and Datadog, including supporting Temporal workers and the infrastructure behind our background processing workflows. You’ll also contribute to the operation of our wider AWS environment as it progressively moves under the Platform team.
This role suits someone who is used to being responsible for live production systems. You’re comfortable responding to incidents, making carefully planned changes and following problems through from service recovery to lasting resolution. Participation in the on-call roster is part of the role.
About You
You may come from platform engineering, site reliability engineering, systems administration or cloud infrastructure. What matters is substantial hands-on experience operating production environments and the judgement that comes with it.
What matters most is how you approach the work:
- You take ownership of the reliability and health of the systems you operate
- You think changes through, including their impact and how to recover if something goes wrong
- You troubleshoot methodically and remain effective under pressure
- You communicate clearly, document useful knowledge and escalate when needed
- You work within agreed technical direction, review and change processes
- You contribute practical improvements and automate recurring manual work
- You’re committed to ongoing learning and secure operational practices
You’ll Be a Strong Fit If You Have
- 5+ years of experience in infrastructure, platform engineering, SRE or a related operational role
- Strong hands-on experience running and troubleshooting Kubernetes in production, ideally AWS EKS
- Strong Linux administration and troubleshooting skills
- Practical AWS experience, including VPCs, IAM, EC2, DNS and networking
- Experience managing infrastructure through Terraform
- Experience operating CI/CD pipelines and supporting production deployments
- Experience investigating production incidents using metrics, logs and monitoring tools
- Scripting skills using Bash, Python or similar
- Experience planning and carrying out production upgrades, maintenance and rollback procedures
Highly Desirable Experience
- GitLab CI/CD and ArgoCD
- Datadog monitoring, logging, dashboards and alerting
- Temporal or similar workflow orchestration platforms, including worker deployments, monitoring and troubleshooting
- EKS Auto Mode, Karpenter and Kubernetes scaling
- Kubernetes networking, ingress, storage and workload configuration
- High-availability systems, capacity management and performance troubleshooting
- AWS security practices and infrastructure access controls
- Supporting containerised applications alongside established Linux and EC2 infrastructure
Why Join Xestro?
- Job security in a long-established, growing healthcare SaaS company
- Meaningful work that directly impacts clinics and patient care
- Supportive, forward-thinking engineering culture with experienced peers
- Hands-on ownership of production infrastructure and opportunities to improve how it operates
- Ongoing learning and a strong focus on secure engineering and operations
- Flexible work arrangements within an onsite-first team
- Sunshine Coast lifestyle with a healthy work-life balance
Australia
New Zealand
United Kingdom
Canada
Singapore
Malaysia





