Senior Infrastructure Technical Lead

  • Australia
  • Melbourne
  • Contract
  • Negotiable

The Role

This is a senior hands-on infrastructure engineering opportunity within a major Australian banking environment, responsible for the engineering, operation and reliability of an enterprise Red Hat OpenShift container platform.
Operating as the L3 technical escalation point, you will resolve complex platform incidents escalated from L1/L2, engineer and automate OpenShift environments, and deliver platform upgrades, security remediation and lifecycle improvements.
The role combines deep OpenShift troubleshooting with platform engineering and automation. You will work across bare-metal and VMware environments, using ArgoCD, Ansible, Helm and Kustomize to maintain consistent, secure and highly available container infrastructure.

Key Responsibilities:

  • Engineer, operate and maintain enterprise OpenShift 4.x clusters across bare-metal and VMware environments, using GitOps and infrastructure-as-code practices.
  • Act as the senior L3 escalation point for complex OpenShift and Kubernetes incidents, leading technical diagnosis, root-cause analysis and permanent problem resolution.
  • Develop and maintain platform automation using ArgoCD, Ansible, Helm and Kustomize, including ownership of GitOps repositories and automated platform configuration.
  • Plan and execute OpenShift platform upgrades, operator upgrades and node patching, while managing cluster capacity, machine sets, machine configuration pools and platform availability.
  • Maintain and improve platform security and reliability across RBAC, SCCs, network policies, secrets, certificates, Vault integration, CVE remediation, monitoring, logging, backup and recovery.
  • Support application teams onboarding to the platform, including namespaces, quotas, access controls and network policies, while providing senior technical guidance on container platform usage.
  • Maintain observability across Prometheus, Grafana, EFK/Loki and Elastic, proactively identifying capacity, performance and reliability issues and improving alert quality.
  • Coordinate platform changes with infrastructure, network, storage and security teams, ensuring appropriate testing, validation and rollback planning.
  • Maintain technical documentation, runbooks and operational procedures, while transferring knowledge to L1/L2 engineers and contributing to continuous platform improvement.

Skills & Experience Required:

  • Deep hands-on experience engineering, operating and troubleshooting Red Hat OpenShift 4.x / Kubernetes platforms within large, complex enterprise environments.
  • Strong understanding of OpenShift internals including operators, networking, ingress, storage, cluster lifecycle, machine configuration, security and troubleshooting.
  • Strong automation and GitOps capability across ArgoCD, Ansible, Helm, Kustomize and Git-based workflows, with experience reducing operational toil through automation.
  • Strong RHEL/Linux and networking knowledge covering TCP/IP, DNS, load balancing, firewalls, proxies and enterprise infrastructure integration.
  • Experience with container platform storage and resilience technologies including OpenShift Data Foundation (ODF)/Ceph, persistent volumes, backup and restore.
  • Strong understanding of container security including RBAC, SCCs, network policies, image security, secrets management, HashiCorp Vault and vulnerability/CVE remediation.
  • Experience with platform monitoring and observability technologies such as Prometheus, Grafana, Elastic, EFK and/or Loki.
  • Scripting capability using Bash and/or Python, combined with strong troubleshooting skills and the ability to work through complex platform issues at L3 level.
  • Experience operating within structured enterprise incident, problem and change management environments, including ServiceNow and major incident processes.

Nice to Have:

  • Experience supporting OpenShift platforms within banking, financial services or similarly large regulated enterprise environments.
  • Exposure to OpenShift EUS-to-EUS upgrades, OpenShift Service Mesh, cert-manager and OADP.
  • Experience working directly with Red Hat support on complex platform issues.
  • Previous responsibility for large-scale, multi-cluster OpenShift estates across both virtualised and bare-metal infrastructure.

What’s in it for You:

  • Initial 12-month contract with a competitive daily rate.
  • Melbourne CBD location with hybrid working arrangements.
  • Work on a large-scale enterprise Red Hat OpenShift container platform within a major Australian banking environment.
  • Highly technical position with genuine ownership across platform engineering, automation, upgrades, security and complex L3 troubleshooting.
  • Opportunity to work across a broad modern container ecosystem including OpenShift, Kubernetes, ArgoCD, Ansible, Helm, Kustomize, ODF/Ceph, Vault and Prometheus/Grafana.

Apply today and Peter Li will reach out to disclose further information.

Apply now

Submit your details and attach your resume below. Hint: make sure all relevant experience is included in your CV and keep your message to the hiring team short and sweet - 2000 characters or less is perfect.