17 Aug
|
Talent International
|
Melbourne
17 Aug
Talent International
Melbourne
This is a senior hands-on infrastructure engineering opportunity within a major Australian banking environment, responsible for the engineering, operation and reliability of an enterprise Red Hat OpenShift container platform .
Operating as the L3 technical escalation point , you will resolve complex platform incidents escalated from L1/L2, engineer and automate OpenShift environments, and deliver platform upgrades, security remediation and lifecycle improvements.
The role combines deep OpenShift troubleshooting with platform engineering and automation. You will work across bare-metal and VMware environments, using ArgoCD, Ansible, Helm and Kustomize to maintain consistent, secure and highly available container infrastructure.
Key Responsibilities:
- Engineer, operate and maintain enterprise OpenShift 4.x clusters across bare-metal and VMware environments, using GitOps and infrastructure-as-code practices.
- Act as the senior L3 escalation point for complex OpenShift and Kubernetes incidents, leading technical diagnosis, root-cause analysis and permanent problem resolution.
- Develop and maintain platform automation using ArgoCD, Ansible, Helm and Kustomize, including ownership of GitOps repositories and automated platform configuration.
- Plan and execute OpenShift platform upgrades, operator upgrades and node patching, while managing cluster capacity, machine sets, machine configuration pools and platform availability.
- Maintain and improve platform security and reliability across RBAC, SCCs, network policies, secrets, certificates, Vault integration, CVE remediation, monitoring, logging, backup and recovery.
- Support application teams onboarding to the platform, including namespaces, quotas, access controls and network policies,
while providing senior technical guidance on container platform usage.
- Maintain observability across Prometheus, Grafana, EFK/Loki and Elastic, proactively identifying capacity, performance and reliability issues and improving alert quality.
- Coordinate platform changes with infrastructure, network, storage and security teams, ensuring appropriate testing, validation and rollback planning.
- Maintain technical documentation, runbooks and operational procedures, while transferring knowledge to L1/L2 engineers and contributing to continuous platform improvement.
Skills & Experience Required:
- Deep hands-on experience engineering, operating and troubleshooting Red Hat OpenShift 4.x / Kubernetes platforms within large, complex enterprise environments.
- Strong understanding of OpenShift internals including operators, networking, ingress, storage, cluster lifecycle, machine configuration, security and troubleshooting.
- Strong automation and GitOps capability across ArgoCD, Ansible, Helm, Kustomize and Git-based workflows, with experience reducing operational toil through automation.
- Strong RHEL/Linux and networking knowledge covering TCP/IP, DNS, load balancing, firewalls, proxies and enterprise infrastructure integration.
- Experience with container platform storage and resilience technologies including OpenShift Data Foundation (ODF)/Ceph, persistent volumes,
backup and restore.
- Strong understanding of container security including RBAC, SCCs, network policies, image security, secrets management, HashiCorp Vault and vulnerability/CVE remediation.
- Experience with platform monitoring and observability technologies such as Prometheus, Grafana, Elastic, EFK and/or Loki.
- Scripting capability using Bash and/or Python, combined with strong troubleshooting skills and the ability to work through complex platform issues at L3 level.
- Experience operating within structured enterprise incident, problem and change management environments, including ServiceNow and major incident processes.
Nice to Have:
- Experience supporting OpenShift platforms within banking, financial services or similarly large regulated enterprise environments.
- Exposure to OpenShift EUS-to-EUS upgrades, OpenShift Service Mesh, cert-manager and OADP.
- Experience working directly with Red Hat support on complex platform issues.
- Previous responsibility for large-scale, multi-cluster OpenShift estates across both virtualised and bare-metal infrastructure.
What’s in it for You:
- Initial 12-month contract with a competitive daily rate.
- Melbourne CBD location with hybrid working arrangements.
- Work on a large-scale enterprise Red Hat OpenShift container platform within a major Australian banking environment.
- Highly technical position with genuine ownership across platform engineering, automation, upgrades, security and complex L3 troubleshooting.
- Prospect to work across a broad modern container ecosystem including OpenShift, Kubernetes, ArgoCD, Ansible, Helm, Kustomize, ODF/Ceph, Vault and Prometheus/Grafana.
#J-18808-Ljbffr
📌 Senior Infrastructure Technical Lead (Melbourne)
🏢 Talent International
📍 Melbourne