DevOps Support Engineer

  • Australia
  • Melbourne
  • Contract
  • Negotiable

This role combines hands-on application and infrastructure support with proactive monitoring, incident response and continuous improvement. You will help maintain the availability, performance and resilience of 24×7 production systems while improving monitoring and early issue detection.

Key responsibilities:

* Monitor and maintain critical OT systems and interfaces
* Triage incidents and perform first-response and recovery activities
* Support production, disaster recovery, test and development environments
* Improve system monitoring, alerting and observability
* Investigate performance issues and security incidents
* Manage application administration, user access and infrastructure activities
* Coordinate system changes, releases, patches, upgrades and certificate renewals
* Analyse system capacity, performance and operational trends
* Support OT communications network monitoring and troubleshooting
* Participate in business continuity and disaster recovery exercises
* Manage technical vendors and service providers
* Contribute to process improvement and emerging OT technology initiatives
* Participate in an on-call support roster

About you:

* Strong application support experience within an IT, OT or network management environment
* Experience supporting highly available 24×7 production systems
* Hands-on knowledge of Linux systems and incident troubleshooting
* Exposure to Kubernetes, AWS, Oracle Database, Nagios or Grafana
* Scripting or programming capability across Python, Bash, SQL or JavaScript
* Experience with change, release and incident management
* Strong analytical, communication and stakeholder engagement skills
* Tertiary qualifications in IT, computer science or a related discipline are preferred

This is an excellent opportunity to work with emerging energy technologies while improving the reliability and performance of systems that support the evolving electricity network.

Apply now or contact Joseph on joseph.petrovski@talentinternational.com for a confidential discussion and further information about the role.

Apply now

Submit your details and attach your resume below. Hint: make sure all relevant experience is included in your CV and keep your message to the hiring team short and sweet - 2000 characters or less is perfect.

Senior Site Reliability Engineer

  • Australia
  • Adelaide
  • Permanent
  • Negotiable

About the Role

We’re looking for an experienced Site Reliability Engineer (SRE) to lead the reliability, performance, and scalability of enterprise platforms. You’ll work closely with software engineering and infrastructure teams to improve automation, cloud platforms, observability, and operational excellence while mentoring other engineers and driving continuous improvement.

Key Responsibilities

  • Design, build and support highly available, scalable, and resilient cloud infrastructure and platforms.
  • Improve system reliability through automation, Infrastructure as Code (IaC), and CI/CD best practices.
  • Lead major incident response, root cause analysis, and implement preventative improvements.
  • Develop and maintain monitoring, logging, alerting, and observability solutions.
  • Establish and manage SLOs, SLIs, and error budgets.
  • Build and maintain CI/CD pipelines and deployment automation.
  • Collaborate with development teams to improve application performance, reliability, and operational readiness.
  • Lead platform modernisation, capacity planning, and disaster recovery initiatives.
  • Provide technical leadership, mentor engineers, and contribute to engineering standards and best practices.
  • Participate in Agile ceremonies and support continuous improvement across the engineering team.

Skills & Experience Required

Essential

  • 7+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, or a similar role.
  • Strong experience administering Windows Server and Linux environments.
  • Experience with cloud platforms such as Azure (preferred), AWS, or GCP.
  • Strong automation and scripting skills using PowerShell, plus Python, Bash, or Go.
  • Experience building and maintaining CI/CD pipelines using Azure DevOps, GitHub Actions, or similar.
  • Hands-on experience with Infrastructure as Code using Terraform, Ansible, or equivalent.
  • Experience supporting Kubernetes and containerised applications.
  • Strong understanding of networking fundamentals including DNS, load balancing, firewalls, and network security.
  • Experience implementing monitoring, logging, and observability platforms.
  • Strong troubleshooting skills with experience leading major incident investigations and root cause analysis.
  • Experience working within Agile and DevOps environments.
  • Excellent communication skills with the ability to work across technical and business stakeholders.

Desirable

  • Relevant tertiary qualification in IT, Computer Science, Engineering, or equivalent experience.
  • Microsoft, Azure, AWS, Kubernetes, Terraform, or ITIL certifications.
  • Experience driving platform modernisation and engineering best practices.
  • Previous experience mentoring engineers or leading technical initiatives.
Apply now

Submit your details and attach your resume below. Hint: make sure all relevant experience is included in your CV and keep your message to the hiring team short and sweet - 2000 characters or less is perfect.