IT Services and IT Consulting👥 5001 employees📍 Auckland, Auckland, NZEst. 1965
Customer focused values. World-class capability.
The right solutions to help you navigate, wherever you are on your journey.
With a breadth of offerings and depth of expertise, there isn’t a s…
Job Overview
We are seeking a motivated and technically capable Site Reliability Engineer to help maintain, and continuously improve the reliability, scalability, security, and performance of our enterprise platforms. You will work across cloud technologies, infrastructure automation, observability, security, and operational processes to ensure services remain resilient and available. This role requires a strong operational mindset, a passion for automation, and the ability to troubleshoot complex technical issues across hybrid and cloud environments.
Key Responsibilities
Reliability Engineering: Implement, and maintain highly available and resilient infrastructure.
Drive continuous service improvements through performance analysis and operational metrics.
Support disaster recovery planning including Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO).
Observability & Monitoring: Develop and maintain monitoring, logging, and alerting solutions.
Proactively identify reliability risks before they impact services.
Analyse system performance and capacity trends.
Create operational dashboards that provide actionable insights.
Incident Response & Problem Management: Participate in and support on-call rotations.
Lead troubleshooting and resolution of production incidents and service outages.
Develop and maintain incident response runbooks and playbooks.
Conduct blameless post-incident reviews and drive corrective actions.
Automation & DevSecOps: Implement and maintain Infrastructure as Code (IaC) solutions.
Automate operational tasks, deployments, and service recovery processes.
Embed security controls throughout the technology lifecycle using DevSecOps principles.
Contribute to CI/CD pipeline development and optimisation.
Cloud & Platform Engineering: Support cloud environments across Azure, AWS and Google Cloud Platform.
Assist with cloud architecture, governance, landing zones, and platform standardisation.
Support containerised workloads using Docker and Kubernetes.
Collaborate with engineering and security teams to improve platform reliability and performance.
Required Qualifications
Ideally, you will bring the following skills and experience to this role:
Technical Experience: 5+ years in Infrastructure, Cloud Engineering, Platform Engineering, Systems Engineering, or Site Reliability Engineering roles.
Experience supporting enterprise cloud platforms including: Microsoft Azure, Amazon Web Services (AWS), Google Cloud Platform (GCP).
Experience administering: Active Directory, Microsoft Entra ID, Microsoft Intune, Google Workspace.
Strong understanding of virtualisation technologies.
Experience with cloud landing zones and platform governance.
Familiarity with containerisation technologies such as Docker and Kubernetes.
Reliability & Operations: Experience implementing monitoring, logging and alerting solutions.
Knowledge of incident management and problem management practices.
Understanding of business continuity, disaster recovery, RTO and RPO requirements.
Security & Governance: Familiarity with: Infrastructure as Code (IaC), DevSecOps practices, Zero Trust security principles, Security and governance frameworks.
Frameworks & Standards: Working knowledge of: NIST Cybersecurity Framework, ITIL, COBIT, ISO 27001, Zero Trust Architecture.
About Datacom
At Datacom you'll be recognised and valued for your contributions. We're growing year on year and can provide stability, career opportunity and a collegial, agile, flat-structured environment that empowers people and promotes autonomy. We care about our people and provide a range of perks such as social events, chill-out spaces, remote working, flexi-hours, professional development courses and other retail discounts to name a few. We operate at the leading edge of technology to help ANZ’s largest enterprise organisations explore possibilities and solve their greatest challenges, so you will never run out of interesting new challenges and opportunities.
Datacom is one of Australia and New Zealand’s largest suppliers of Information Technology professional services. We have managed to maintain a dynamic, agile, small business feel that is often diluted in larger organisations of our size. It's our people that give Datacom its unique culture and energy that you can feel from the moment you meet with us.
We care about our people and provide a range of perks such as social events, chill-out spaces, remote working, flexi-hours and professional development courses to name a few. You’ll have the opportunity to learn, develop your career, connect and bring your true self to work. You will be recognised and valued for your contributions and be able to do your work in a collegial, flat-structured environment.
We operate at the forefront of technology to help Australia and New Zealand’s largest enterprise organisations explore possibilities and solve their greatest challenges, so you will never run out of interesting new challenges and opportunities.
We want Datacom to be an inclusive and welcoming workplace for everyone.
Please let Datacom know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🇪🇺 Secure Your EU Traffic
Ensure digital sovereignty for your infrastructure. Get EU static IPs with full data residency for compliance and peace of mind.